Which company has the best AI model end of September?
Anthropic is still the likeliest name to be on top at the end of September, but the market looks a bit too confident given how quickly leaderboard positions can change. I would price Yes at 91%, with the main uncertainty coming from a late-month jump by OpenAI or Google.
Analysis
Anthropic entering late September with the market already pricing a very high probability suggests that the company is either currently leading or very close to the top of the Arena leaderboard. In a ranking system like this, the incumbent advantage is meaningful because the leader is often protected by strong performance across a broad mix of prompts, and challengers need a clear quality edge to overtake it. If Anthropic’s current model is already first or near first, it does not need to improve much to hold the spot through month-end, which supports a high Yes probability.
That said, the market price near 95% is so elevated that it likely assumes not just current leadership but also a relatively low chance of disruption over the next couple of weeks. That is a lot to ask in a fast-moving model race. A single strong release, a leaderboard adjustment, or a model update from a large competitor can move these rankings quickly, especially when the gap between top systems is often narrow. Because the market resolves strictly on the leaderboard position at a fixed timestamp, even a brief late-month reversal would be enough to flip the outcome.
From a historical and competitive standpoint, Anthropic is one of the most credible candidates to hold the top spot because its frontier models have been consistently competitive in chat-style and general reasoning evaluations. However, OpenAI and Google are the most plausible threats, and both have the resources and incentive to make late pushes. The key question is less whether Anthropic is among the very best models and more whether it can remain the single top-ranked model at the exact check time, which is inherently harder and makes a modest discount to the market price appropriate.
Overall, the evidence still favors Yes by a wide margin, but not enough to justify the near-certain market level. The combination of current strength, brand momentum, and likely existing leadership supports Anthropic, while the remaining risk is a fast late-month leaderboard shift from another frontier lab. That leaves me comfortably bullish on Yes, but below the current market-implied confidence.
Arguments
For
- Arguments for Yes: Anthropic has been one of the most consistently competitive frontier AI companies, so it has a strong chance to remain first if it is already near the top.
- Arguments for Yes: The market’s heavy Yes bias implies the current leaderboard state likely already favors Anthropic, and incumbency often persists over short horizons.
Against
- Arguments against Yes: The market is pricing an extremely high probability, leaving little room for any surprise from a rival model release.
- Arguments against Yes: OpenAI and Google have the technical and product capacity to leapfrog Anthropic quickly if they push a new or improved model before the check date.
Key drivers
- Anthropic likely starts from a strong leaderboard position, which is a meaningful advantage in a short time window.
- The most realistic path to No is a late-month update from OpenAI or Google that nudges one of their models above Anthropic.
Risk factors
- A new model release or refresh near the end of September could quickly overturn the current ranking.
- Leaderboard ranks can be sensitive to small score differences, making the exact check time unusually important.
Scenarios
Best case
Anthropic keeps the first-place rank through September 30, benefiting from stable performance while competitors fail to surpass its score.
Most likely
Anthropic remains one of the top models and is either first already or reclaims first by month-end, but the race stays close enough that a late upset remains the main risk.
Worst case
A late update from a rival model pushes Anthropic into second place or lower right before the check time, causing a No resolution.
More from this day
- economyPolymarket3mo
How many Fed rate cuts in 2026?
AI11%MKT93%Edge-82HypedI think there is still a meaningful chance of at least one Fed cut in 2026, but the market is pricing in a very strong hold case for a reason. My estimate is that there is about an 11% chance of no rate cuts at all in 2026.
- techPolymarket3mo
How many more Millennium Prize Problems will AI solve in 2026?
AI88%MKT19%Edge+69Hidden GemI think there is a strong chance that no qualifying Millennium Prize Problems will be counted for 2026 under this market’s rules, even though OpenAI’s Navier-Stokes claim makes the broader “AI solved a Millennium Problem” narrative look more plausible. The exclusion of Navier-Stokes from resolution is the key reason the zero-solved outcome still looks likelier than not.
- FinancialsKalshi13y
Will OpenAI or Anthropic IPO first?
AI34%MKT95%Edge-61HypedMy independent view is that OpenAI is less likely than the market implies to be the first of the two to IPO. The current reporting tilt favors Anthropic moving earlier, and OpenAI’s process appears more valuation- and timing-constrained.