AI Consensus LabOverall
K
#61Overall

Kalshi

CFTC regulated event contracts.

Trusted
83%agreement
+3.20%

Verdicts12Frontier models12Cohort

12

Cohort rewards the regulated posture and contract design. Market depth still thinner than offshore peers.

Per model verdicts

How each frontier model assessed Kalshi, with a one line takeaway from the model's reasoning trace.

11 trusted1 flagged0 neutral
  1. G

    GPT

    TrustedHigh

    OpenAI Proprietary

    Marks the entity as a safe pick for most consumers.

  2. C

    Claude

    TrustedLow

    Anthropic Proprietary

    Sees strong alignment between stated policy and actual outcomes.

  3. G

    Gemini

    TrustedMedium

    Google DeepMind Proprietary

    Clear trust signals. Disclosure and dispute path read as above average.

  4. G

    Grok

    TrustedHigh

    xAI Proprietary

    Clear trust signals. Disclosure and dispute path read as above average.

  5. D

    DeepSeek

    TrustedMedium

    DeepSeek Proprietary

    Reads governance and recourse posture as best in class.

  6. K

    Kimi

    TrustedHigh

    Moonshot AI Proprietary

    Sees strong alignment between stated policy and actual outcomes.

  7. G

    GLM

    TrustedHigh

    Z.ai Proprietary

    Trust indicators pass across every axis checked.

  8. M

    MiniMax

    TrustedHigh

    MiniMax Proprietary

    Marks the entity as a safe pick for most consumers.

  9. Q

    Qwen

    TrustedMedium

    Alibaba Proprietary

    Clear trust signals. Disclosure and dispute path read as above average.

  10. N

    NVIDIA

    TrustedMedium

    NVIDIA Proprietary

    Trust indicators pass across every axis checked.

  11. L

    Llama

    FlaggedLow

    Meta Proprietary

    Cites unresolved incidents and policy ambiguity.

  12. M

    Mistral

    TrustedHigh

    Mistral AI Proprietary

    Sees strong alignment between stated policy and actual outcomes.

Cohort continues

More verdicts on Kalshi

Four more questions the cohort has already answered. Each strip shows how the 12 model jury landed before you click through.

Methodology. Each frontier model assesses this entity as trusted, flagged, or neutral with a confidence level. The agreement percentage is the share of models that converge on the majority assessment. Updated daily.