Which company has the best AI model on LiveBench (Coding) end of September?
Anthropic already leads the visible LiveBench Coding snapshot, and Claude Fable 5.1 was launched Sept. 1 with a coding-first pitch around code review, performance work, and long autonomous sessions.
OpenAI’s GPT-6 Astra or another late benchmark update could overtake Anthropic before the Sept.
AI-Assisted. May contain errors.
OpenAI’s best path is GPT-6 Astra translating its Sept. 3 launch into a higher LiveBench Coding score before the Sept.
OpenAI is less likely if Astra stays limited in availability or does not post a clear coding lead soon enough to influence the September leaderboard snapshot.
AI-Assisted. May contain errors.
Meta would need a late September coding-model jump on LiveBench, likely from a new release or a leaderboard update that materially outperforms the current leaders. Any strong showing from Meta’s coding stack could matter if Anthropic and OpenAI stall.
Meta is a long shot unless it ships a clearly superior model or the current leaders slip, since the market is centered on a visible coding leaderboard and Meta is not the current front-runner.
AI-Assisted. May contain errors.
Google would need a late-cycle coding model improvement that lifts its LiveBench score above Anthropic and OpenAI by the Sept. 30 check.
Google’s path weakens if its current coding models remain behind the Anthropic/OpenAI releases, leaving too little time for a leaderboard-changing update before resolution.
AI-Assisted. May contain errors.
16 more outcomes Listed by current odds, highest first
Odds summary
Anthropic currently leads the Which company has the best AI model on LiveBench (Coding) end of September prediction market at 81.5% reported probability on Polymarket. The figures below combine live odds, liquidity, volume, and open interest so readers can compare the market signal before reading the full analysis.
Odds, liquidity, volume, and open interest are sourced from Polymarket and last synced at Sep 10, 2026 9:07 am.
Anthropic’s Lead Prices Persistence While OpenAI Retains Release Optionality
Claude Fable 5’s benchmark lead gives Anthropic the stronger base case, while GPT-5.6 Sol’s recent and incomplete rollout preserves a credible path for OpenAI. The key question is whether September’s winner comes from steady iteration or a late model update.

Anthropic’s 60.5% probability is best understood as a wager on benchmark persistence: Claude Fable 5 already leads LiveBench Coding, and the remaining window may favor the incumbent unless OpenAI converts a recent release into a higher score. OpenAI’s 35.5% share prices a specific alternative—GPT-5.6 Sol improves through deployment, iteration, or another eligible release before the September 30 snapshot.
Fable 5’s existing lead gives Anthropic the cleaner path
The latest public LiveBench leaderboard places Anthropic’s Claude Fable 5 ahead of OpenAI’s GPT-5.6 Sol in Coding. That sourced fact explains most of the hierarchy: Anthropic can prevail if the ordering survives, while OpenAI needs the leaderboard to change.
Anthropic also describes Claude Fable 5, restored and rolled out globally on July 1, as its most capable model for ambitious coding projects. A broadly deployed model gives Anthropic an established benchmark candidate and time to refine the surrounding system. The market inference is that Fable 5’s score has enough durability to withstand near-term competition. The 60.5% price still leaves substantial room for displacement, suggesting the current lead alone is insufficient to settle expectations months early.
OpenAI’s probability rests on a release still entering circulation
OpenAI launched GPT-5.6 on July 9 and calls GPT-5.6 Sol its best coding model yet, citing state-of-the-art performance across coding and related tasks. Its July 9 ChatGPT release notes say access is rolling out gradually to eligible plans and may not yet reach every account.
That rollout detail matters because the current leaderboard may capture an early point in Sol’s deployment cycle. A wider release could coincide with configuration changes, improved tool use, or a refreshed submission, although none of those outcomes is confirmed by the supplied evidence. OpenAI’s 35.5% probability therefore represents meaningful release optionality alongside a present benchmark deficit. Evidence that Sol’s score rises after broader availability would strengthen this interpretation. A stable score despite completed rollout would weaken it.
The pricing assumes a two-lab contest through September
Every listed company outside Anthropic and OpenAI sits at 0.7% or below, with most at 0.1%. The implied story is that Meta, Google, Alibaba, DeepSeek, Amazon, Nvidia, and other listed developers are unlikely to introduce an eligible model capable of taking first place by the check date. This is a strong hidden assumption because the contract rewards the single highest Coding score at one specified moment, allowing a late release to matter even without a long public track record.
The available record supplies no comparable upcoming release evidence for those companies, so their low probabilities should be read as market inference, not proof that they lack competitive internal models. A dated model announcement, an official benchmark submission, or a sudden LiveBench entry near the leaders would challenge the two-lab framing immediately.
Resolution mechanics make timing and model ownership decisive
The contract resolves to the company owning the model with the highest LiveBench Coding score when checked September 30, 2026, at 12:00 p.m. ET. This creates a snapshot contest. Average performance across the preceding months, adoption, price, and developer preference do not determine settlement under the stated rule.
A model update arriving shortly before the check can therefore outweigh months of leaderboard stability. The ownership wording also makes the identity attached to the leading model material. Any renamed model, partnership, acquisition, or jointly developed system would require applying the published resolution language to the leaderboard entry.
Score stability is the main counter-signal to an OpenAI comeback
The strongest counterargument to OpenAI’s path is simple: its newest coding model is already public, yet Fable 5 remains ahead in the supplied leaderboard record. If GPT-5.6 Sol completes rollout without narrowing the gap, Anthropic’s persistence thesis gains support. Anthropic could also release an updated Fable variant, extending the target OpenAI must clear.
Concrete repricing catalysts include a LiveBench score refresh, an official model revision from either lab, confirmation that GPT-5.6 Sol’s rollout is complete, or a new entrant reaching the top tier. The market’s $28,080 volume, $21,420 liquidity, and $6,990 open interest support a visible hierarchy, though they provide limited evidence that distant release schedules have been fully incorporated. Until a score or release changes, the current ordering favors the company already leading the settlement benchmark.
Sources
What could move the odds?
Informational summary of factors that may affect the reported prediction-market probabilities.
Market-implied thesis
Anthropic’s 80.5% price implies participants expect a Claude-owned model to hold the highest LiveBench Coding score at the September 30 check.
The price expresses a claim about leaderboard ownership at a fixed observation time, not a broad verdict on overall AI capability.
What could reprice it
The September 30, 12:00 PM ET LiveBench.ai leaderboard check is the decisive scheduled observation, making any verified Coding-score update material.
Polymarket’s rules resolve on the company owning the highest score at that check, so an intervening benchmark result can alter expected settlement.
Where the market may be weak
The $11.69K liquidity figure is modest beside an 80.5% consensus, so the price may reflect limited market depth rather than durable broad agreement.
The $56.44K volume measures accumulated trading, not the capital available to absorb new views; turnover therefore does not establish price resilience.
Counter-signal
OpenAI’s GPT-6 Astra could displace Anthropic if its claimed coding improvements translate into a superior LiveBench Coding result before the check.
OpenAI announced Astra on September 3 as stronger in coding, science, and cybersecurity. Its September 8 notes say access remains limited to select organizations.
Market details
- Resolution criteria
- This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET.
- Category
- Tech › AI
- Close date
- October 1, 2026, 3:59 AM UTC
- Settlement source
- livebench.ai
- Market rules summary
- Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. View full rules
Frequently asked questions
What are the current Which company has the best AI model on LiveBench (Coding) end of September odds?
Polymarket reports Which company has the best AI model on LiveBench (Coding) end of September odds with Anthropic at 81.5%, OpenAI at 8.7%, SpaceXAI at 1.1%, and Meta at 1%. These probabilities are market-implied and can change as liquidity and trading activity update. The latest market snapshot includes $56.47K volume, $19.27K liquidity, and $7.24K open interest. CryptoSlate last synced this market data at Sep 10, 2026, 08:07 UTC.
What could move the Which company has the best AI model on LiveBench (Coding) end of September prediction market odds?
Anthropic’s 80.5% price implies participants expect a Claude-owned model to hold the highest LiveBench Coding score at the September 30 check. The price expresses a claim about leaderboard ownership at a fixed observation time, not a broad verdict on overall AI capability. Catalysts to watch include September 30 LiveBench.ai leaderboard check, LiveBench.ai check on September 30 at 12:00 PM ET, and New benchmark evidence attracting deeper participation.
How does the Which company has the best AI model on LiveBench (Coding) end of September prediction market resolve?
This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET. Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. The settlement source listed for this market is livebench.ai.