Third-best Code Arena WebDev AI Lab end of August?
Moonshot has a real path to finishing in third, but the leaderboard is competitive enough that the exact third-place slot looks less likely than the market price implies. I estimate about a 27% chance that Moonshot is the third-best WebDev lab at the August 31 check.
Analysis
The key point is that this market does not ask whether Moonshot is strong overall, but whether Moonshot lands in the exact third position among labs at a single noon ET snapshot on August 31. The recent context suggests Moonshot is already near the top tier, with Kimi K3 Max appearing around second in a recent live view, while other Moonshot models sit much lower and do not change the company’s standing on their own. That makes the outcome highly dependent on the relative position of Moonshot’s best model versus the best models from Anthropic, Qwen, OpenAI, and GLM rather than on Moonshot’s broader model count.
The competitive backdrop argues against a high probability of exactly third. When several labs are clustered with small Elo gaps, a one-step shuffle can happen, but the most common outcomes are usually that the leader holds, a challenger overtakes the leader, or a lab slips more than one position rather than landing cleanly in the middle. If Moonshot is currently second, then yes requires at least one rival to pass it by the deadline while Moonshot still stays ahead of the rest, which is plausible but still a fairly specific configuration. If Moonshot is actually already third in the underlying lab ranking, that would help the yes case, but the provided evidence points more to Moonshot being near the top than being safely locked into third.
The late-August timing adds some uncertainty because new releases, leaderboard recalculations, or modest score updates can move ranks quickly near the cutoff. That helps the yes side because a close rival could edge Moonshot down one place without a dramatic collapse, but it also hurts the yes side because Moonshot could just as easily strengthen or remain stable and finish second instead of third. The current market price around 30% is reasonable as a reference point, but my independent read is slightly lower because exact third is a narrow landing zone and Moonshot’s strongest visible models appear more likely to remain top-two material than to settle precisely into third.
Arguments
For
- Arguments for Yes: The top of the leaderboard is tightly packed, so a modest improvement by a rival could push Moonshot down exactly one spot.
- Arguments for Yes: The noon ET resolution time creates a single snapshot that can capture a temporary rank shuffle in Moonshot’s favor.
Against
- Arguments against Yes: Current evidence suggests Moonshot is already near second place, which makes third an awkward middle outcome rather than the base case.
- Arguments against Yes: If Moonshot’s flagship model stays strong, the more likely result is holding second instead of slipping to third.
Key drivers
- Moonshot already appears close to the top of the WebDev board, so a small relative change could move it into third.
- The leaderboard has several neighboring labs with small score gaps, making one-position shuffles plausible before the cutoff.
- The market resolves on a single noon snapshot, so a brief late-month rank swing would be enough to decide the outcome.
- Moonshot’s outcome depends on its best lab entry versus competitors’ best entries, not on the weaker Moonshot models lower on the board.
Risk factors
- Moonshot may simply remain second if its flagship model stays stable while rivals fail to overtake it.
- Moonshot could fall below third if another lab releases or updates a stronger model late in the month.
- Leaderboard updates can be noisy, which means a temporary dip or rebound may not persist until the resolution time.
- Exact-third outcomes are fragile because a small change can shift Moonshot from second to fourth rather than landing neatly in third.
Scenarios
Best case
A close competitor from Anthropic, Qwen, OpenAI, or GLM edges ahead of Moonshot by the end of August, while Moonshot remains ahead of everyone else below, producing an exact third-place finish.
Most likely
Moonshot remains in the top tier but lands in either second or fourth rather than precisely third, with late leaderboard movement deciding which side of the narrow boundary it ends up on.
Worst case
Moonshot either keeps second place through the cutoff or drops more than one spot because another lab overtakes it decisively, leaving Moonshot outside third.
More from this day
- CompaniesKalshi1y
Starbucks total global stores in 2026
AI89%MKT9%Edge+80Hidden GemStarbucks looks likely to finish 2026 above 41,800 stores. The company’s reported Q3 base of 41,304 and full-year guidance for 600 to 650 net new coffeehouses leave a meaningful cushion over the threshold.
- PoliticsKalshi1y
2026: Trump's bad year?
AI73%MKT11%Edge+62Hidden GemTrump looks materially more likely than not to have a genuinely adverse 2026, with legal exposure, court fights, and internal political resistance creating several paths to a bad year. The market’s 11% yes price looks far too low unless the event is defined very narrowly.
- FinancialsKalshi13y
Will OpenAI or Anthropic IPO first?
AI30%MKT82%Edge-52HypedI estimate OpenAI has only about a 30% chance of beating Anthropic to the IPO. Anthropic’s reported late-2026 target and OpenAI’s apparent drift toward 2027 make Anthropic the likelier first mover.