Which company has the best AI model end of September?
Anthropic is still a strong favorite to lead this specific text-arena leaderboard at the end of September, but the current price looks somewhat aggressive given how quickly top models can shuffle. I would price Yes below the market, though still as the most likely outcome.
Analysis
Anthropic enters this market with a meaningful structural advantage: it has repeatedly been one of the strongest performers in chat-style and instruction-following benchmarks, and this market resolves on a single overall text arena ranking rather than broad product adoption or revenue. That matters because a company can stay near the top of a leaderboard with a model that is exceptionally good at human preference, coherence, and day-to-day conversational quality even if another firm has the more technically impressive or multimodal system. The market price already reflects that Anthropic is often viewed as a leader in this class of evaluation, so the main question is not whether Anthropic can be competitive, but whether it can remain first when the table is checked at a specific moment at the end of the month.
The strongest argument for Anthropic is path dependence. Companies that are already leading a text arena often have a high chance of retaining the top spot for some period because incremental improvements, rapid refresh cycles, and favorable alignment with the judge population can preserve rank even as competitors close the gap. If Anthropic continues to ship steady model improvements while rivals are focused on broader releases, it can maintain enough of an edge to stay ahead in a leaderboard that likely rewards polished conversational performance. The size of the event volume also suggests a relatively efficient market, so the high Yes price is not just speculation; it likely incorporates informed expectations that Anthropic has a real edge in this exact setting.
The main reason to be less bullish than the market is that a 87.5% implied probability leaves very little room for uncertainty, and this kind of ranking is inherently fragile. At the end of a month, one strong release from OpenAI, Google, xAI, or another lab could easily displace Anthropic, especially if the arena has recently begun to favor a new generation of models or if the scoring gap between the top few entries is thin. Because the market resolves on a snapshot rather than an average over time, even a brief leader change on the exact check date would flip the outcome. My assessment is that Anthropic is still the likeliest single company to be first, but the probability of a late-month reshuffle is high enough that the true chance is meaningfully below the market-implied level.
Arguments
For
- Arguments for Yes: Anthropic has repeatedly shown strength in instruction-following and conversational quality, which tends to translate well into arena rankings.
- Arguments for Yes: If the current leader is already Anthropic, incumbency increases the odds that it remains first through the end-of-month snapshot.
Against
- Arguments against Yes: The market price assumes Anthropic is almost certain to stay on top, but frontier model rankings often change with little warning.
- Arguments against Yes: A single strong release from a rival company could push Anthropic to second without needing a broad or lasting advantage.
Key drivers
- Anthropic has a strong historical fit for human preference and chat-based evaluation, which is exactly what this leaderboard measures.
- The resolution is a single snapshot at month-end, so a late rival model release can overturn a lead very quickly.
Risk factors
- A competitor could launch a materially better model before September 30 and immediately take first place.
- Small score differences at the top of arena-style rankings can make the outcome highly sensitive to minor leaderboard movements.
Scenarios
Best case
Anthropic maintains or extends its lead through September, while competitors either do not release a stronger model or fail to displace it on the specific arena metric used for resolution.
Most likely
Anthropic remains near the top of the leaderboard and has a legitimate shot at first, but the exact month-end snapshot is uncertain enough that the outcome is not close to guaranteed.
Worst case
Another company ships a better text model before the check time and takes first place on the leaderboard, leaving Anthropic in second or lower.
More from this day
- PoliticsKalshi2y
Which agencies will Trump eliminate?
AI89%MKT26%Edge+63Hidden GemUSAID looks very likely to count as eliminated during Trump’s term. The reporting provided strongly suggests the agency was dismantled early, and the market price appears far too low unless the contract definition requires a formal statutory repeal that has not occurred.
- CompaniesKalshi1y
Starbucks total global stores in 2026
AI74%MKT11%Edge+63Hidden GemStarbucks is very likely to clear 41,800 global stores sometime in 2026. The market appears to be pricing in a much slower store-opening cadence than Starbucks has historically maintained.
- pop culturePolymarketTomorrow
"Resident Evil" Opening Weekend Box Office
AI82%MKT26%Edge+56Hidden GemResident Evil is more likely than not to open below 50 million domestically. That threshold is high enough that only a genuinely breakout event release would clear it, and the franchise’s historical domestic openings have usually been well under that level.