Which company has the best AI model end of September?
Anthropic is still the likeliest name to be on top at the end of September, but the market looks a bit too confident given how quickly leaderboard positions can change. I would price Yes at 91%, with the main uncertainty coming from a late-month jump by OpenAI or Google.
Analysis
Anthropic entering late September with the market already pricing a very high probability suggests that the company is either currently leading or very close to the top of the Arena leaderboard. In a ranking system like this, the incumbent advantage is meaningful because the leader is often protected by strong performance across a broad mix of prompts, and challengers need a clear quality edge to overtake it. If Anthropic’s current model is already first or near first, it does not need to improve much to hold the spot through month-end, which supports a high Yes probability.
That said, the market price near 95% is so elevated that it likely assumes not just current leadership but also a relatively low chance of disruption over the next couple of weeks. That is a lot to ask in a fast-moving model race. A single strong release, a leaderboard adjustment, or a model update from a large competitor can move these rankings quickly, especially when the gap between top systems is often narrow. Because the market resolves strictly on the leaderboard position at a fixed timestamp, even a brief late-month reversal would be enough to flip the outcome.
From a historical and competitive standpoint, Anthropic is one of the most credible candidates to hold the top spot because its frontier models have been consistently competitive in chat-style and general reasoning evaluations. However, OpenAI and Google are the most plausible threats, and both have the resources and incentive to make late pushes. The key question is less whether Anthropic is among the very best models and more whether it can remain the single top-ranked model at the exact check time, which is inherently harder and makes a modest discount to the market price appropriate.
Overall, the evidence still favors Yes by a wide margin, but not enough to justify the near-certain market level. The combination of current strength, brand momentum, and likely existing leadership supports Anthropic, while the remaining risk is a fast late-month leaderboard shift from another frontier lab. That leaves me comfortably bullish on Yes, but below the current market-implied confidence.
Arguments
For
- Arguments for Yes: Anthropic has been one of the most consistently competitive frontier AI companies, so it has a strong chance to remain first if it is already near the top.
- Arguments for Yes: The market’s heavy Yes bias implies the current leaderboard state likely already favors Anthropic, and incumbency often persists over short horizons.
Against
- Arguments against Yes: The market is pricing an extremely high probability, leaving little room for any surprise from a rival model release.
- Arguments against Yes: OpenAI and Google have the technical and product capacity to leapfrog Anthropic quickly if they push a new or improved model before the check date.
Key drivers
- Anthropic likely starts from a strong leaderboard position, which is a meaningful advantage in a short time window.
- The most realistic path to No is a late-month update from OpenAI or Google that nudges one of their models above Anthropic.
Risk factors
- A new model release or refresh near the end of September could quickly overturn the current ranking.
- Leaderboard ranks can be sensitive to small score differences, making the exact check time unusually important.
Scenarios
Best case
Anthropic keeps the first-place rank through September 30, benefiting from stable performance while competitors fail to surpass its score.
Most likely
Anthropic remains one of the top models and is either first already or reclaims first by month-end, but the race stays close enough that a late upset remains the main risk.
Worst case
A late update from a rival model pushes Anthropic into second place or lower right before the check time, causing a No resolution.
More from this day
- pop culturePolymarket11d
"Spider-Man: Brand New Day" total domestic gross by September 30?
AI99%MKT2%Edge+97Hidden GemSpider-Man: Brand New Day finishing below 940m domestic by September 30 looks overwhelmingly likely. The threshold is so high that only an extreme, record-breaking performance would threaten it.
- pop culturePolymarket3mo
Where will 2026 rank among the hottest years on record?
AI18%MKT75%Edge-57Hyped2026 is more likely than not to fall short of being the single hottest year on record. The market appears to be pricing in a continuation of extreme global warmth, but the odds still favor at least one other year, especially 2024 or 2025, finishing hotter on the NASA index.
- PoliticsKalshi2y
Who will Trump pardon?
AI4%MKT51%Edge-47HypedI estimate Barron Trump’s chance of receiving a presidential pardon before January 21, 2029 at very low, around 4%. Absent a specific federal criminal exposure or public controversy, there is little reason to expect a pardon for a private family member.