Which company has the best AI model on LiveBench (Coding) end of September?
Anthropic is the current LiveBench Coding leader, so a hold or further gain in the next leaderboard refreshes would keep it on top through the September 30 check. Claude Sonnet 5’s rollout and any follow-on coding-tuned updates could reinforce that edge if rivals do not overtake it.
A stronger OpenAI release or a LiveBench refresh that lifts another lab above Claude would break Anthropic’s lead before the September 30 snapshot.
AI-Assisted. May contain errors.
OpenAI has an active coding roadmap, including GPT-5.3-Codex and the newer GPT-5.6 Sol positioning, so a broader or stronger release could lift its LiveBench Coding score quickly. If that model ships before September 30 and outperforms Claude, the leaderboard can flip.
If OpenAI’s next coding model stays limited, slips past the deadline, or underperforms Anthropic on LiveBench, it will remain behind the current leader.
AI-Assisted. May contain errors.
Meta would need a late coding-model jump that outperforms the current front-runners on LiveBench, likely via a new release or a major refresh of its Llama line. Any benchmark improvement would have to arrive before the September 30 leaderboard check and beat Anthropic and OpenAI directly.
Without a standout coding release, Meta is likely to remain behind the current leaders and never reach the top of the LiveBench Coding board.
AI-Assisted. May contain errors.
16 more outcomes Listed by current odds, highest first
Odds summary
Anthropic currently leads the Which company has the best AI model on LiveBench (Coding) end of September prediction market at 54.5% reported probability on Polymarket. The figures below combine live odds, liquidity, volume, and open interest so readers can compare the market signal before reading the full analysis.
Odds, liquidity, volume, and open interest are sourced from Polymarket and last synced at Aug 7, 2026 11:47 pm.
Anthropic’s Lead Prices Persistence While OpenAI Retains Release Optionality
Claude Fable 5’s benchmark lead gives Anthropic the stronger base case, while GPT-5.6 Sol’s recent and incomplete rollout preserves a credible path for OpenAI. The key question is whether September’s winner comes from steady iteration or a late model update.

Anthropic’s 60.5% probability is best understood as a wager on benchmark persistence: Claude Fable 5 already leads LiveBench Coding, and the remaining window may favor the incumbent unless OpenAI converts a recent release into a higher score. OpenAI’s 35.5% share prices a specific alternative—GPT-5.6 Sol improves through deployment, iteration, or another eligible release before the September 30 snapshot.
Fable 5’s existing lead gives Anthropic the cleaner path
The latest public LiveBench leaderboard places Anthropic’s Claude Fable 5 ahead of OpenAI’s GPT-5.6 Sol in Coding. That sourced fact explains most of the hierarchy: Anthropic can prevail if the ordering survives, while OpenAI needs the leaderboard to change.
Anthropic also describes Claude Fable 5, restored and rolled out globally on July 1, as its most capable model for ambitious coding projects. A broadly deployed model gives Anthropic an established benchmark candidate and time to refine the surrounding system. The market inference is that Fable 5’s score has enough durability to withstand near-term competition. The 60.5% price still leaves substantial room for displacement, suggesting the current lead alone is insufficient to settle expectations months early.
OpenAI’s probability rests on a release still entering circulation
OpenAI launched GPT-5.6 on July 9 and calls GPT-5.6 Sol its best coding model yet, citing state-of-the-art performance across coding and related tasks. Its July 9 ChatGPT release notes say access is rolling out gradually to eligible plans and may not yet reach every account.
That rollout detail matters because the current leaderboard may capture an early point in Sol’s deployment cycle. A wider release could coincide with configuration changes, improved tool use, or a refreshed submission, although none of those outcomes is confirmed by the supplied evidence. OpenAI’s 35.5% probability therefore represents meaningful release optionality alongside a present benchmark deficit. Evidence that Sol’s score rises after broader availability would strengthen this interpretation. A stable score despite completed rollout would weaken it.
The pricing assumes a two-lab contest through September
Every listed company outside Anthropic and OpenAI sits at 0.7% or below, with most at 0.1%. The implied story is that Meta, Google, Alibaba, DeepSeek, Amazon, Nvidia, and other listed developers are unlikely to introduce an eligible model capable of taking first place by the check date. This is a strong hidden assumption because the contract rewards the single highest Coding score at one specified moment, allowing a late release to matter even without a long public track record.
The available record supplies no comparable upcoming release evidence for those companies, so their low probabilities should be read as market inference, not proof that they lack competitive internal models. A dated model announcement, an official benchmark submission, or a sudden LiveBench entry near the leaders would challenge the two-lab framing immediately.
Resolution mechanics make timing and model ownership decisive
The contract resolves to the company owning the model with the highest LiveBench Coding score when checked September 30, 2026, at 12:00 p.m. ET. This creates a snapshot contest. Average performance across the preceding months, adoption, price, and developer preference do not determine settlement under the stated rule.
A model update arriving shortly before the check can therefore outweigh months of leaderboard stability. The ownership wording also makes the identity attached to the leading model material. Any renamed model, partnership, acquisition, or jointly developed system would require applying the published resolution language to the leaderboard entry.
Score stability is the main counter-signal to an OpenAI comeback
The strongest counterargument to OpenAI’s path is simple: its newest coding model is already public, yet Fable 5 remains ahead in the supplied leaderboard record. If GPT-5.6 Sol completes rollout without narrowing the gap, Anthropic’s persistence thesis gains support. Anthropic could also release an updated Fable variant, extending the target OpenAI must clear.
Concrete repricing catalysts include a LiveBench score refresh, an official model revision from either lab, confirmation that GPT-5.6 Sol’s rollout is complete, or a new entrant reaching the top tier. The market’s $28,080 volume, $21,420 liquidity, and $6,990 open interest support a visible hierarchy, though they provide limited evidence that distant release schedules have been fully incorporated. Until a score or release changes, the current ordering favors the company already leading the settlement benchmark.
Sources
What could move the odds?
Informational summary of factors that may affect the reported prediction-market probabilities.
Market-implied thesis
Market pricing implies Anthropic is likelier than any single rival to own the model leading LiveBench Coding at the specified check.
This is a relative winner claim, not a claim that Anthropic will retain a permanent coding advantage. The research context identifies Anthropic as the current leader in the latest release.
What could reprice it
A stronger coding-model release or a LiveBench leaderboard refresh before the September check is the clearest route to material repricing.
OpenAI describes GPT-5.6 Sol as stronger in coding, while its prior GPT-5.3-Codex release shows continuing coding-model iteration. Neither source supplies a dated release or benchmark-refresh schedule.
Where the market may be weak
The deciding leaderboard snapshot precedes the listed close time, leaving unclear whether trading can continue after the observation that fixes settlement.
Rules set the leaderboard check for September 30 at 12:00 PM ET, while the listed close is September 30 at 11:59 PM UTC. That timing gap may complicate interpretation of post-snapshot price discovery.
Counter-signal
OpenAI could displace Anthropic if GPT-5.6 Sol is broadly exposed and records the highest LiveBench Coding score by the check.
OpenAI says GPT-5.6 Sol has stronger coding capabilities, and its February 5 GPT-5.3-Codex launch was described as its most capable agentic coding model. The supplied sources do not establish a LiveBench score for either model.
Market details
- Resolution criteria
- This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET.
- Category
- Tech › AI
- Close date
- September 30, 2026, 11:59 PM UTC
- Settlement source
- livebench.ai
- Market rules summary
- Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. View full rules
Frequently asked questions
What are the current Which company has the best AI model on LiveBench (Coding) end of September odds?
Polymarket reports Which company has the best AI model on LiveBench (Coding) end of September odds with Anthropic at 54.5%, OpenAI at 30.5%, Meta at 4.9%, and SpaceXAI at 0.7%. These probabilities are market-implied and can change as liquidity and trading activity update. The latest market snapshot includes $29.83K volume, $12.46K liquidity, and $6.64K open interest. CryptoSlate last synced this market data at Aug 7, 2026, 22:47 UTC.
What could move the Which company has the best AI model on LiveBench (Coding) end of September prediction market odds?
Market pricing implies Anthropic is likelier than any single rival to own the model leading LiveBench Coding at the specified check. This is a relative winner claim, not a claim that Anthropic will retain a permanent coding advantage. The research context identifies Anthropic as the current leader in the latest release. Catalysts to watch include Model releases or a leaderboard refresh before September 30, Broader GPT-5.6 Sol availability or a LiveBench refresh, and Clarification of trading and resolution timing.
How does the Which company has the best AI model on LiveBench (Coding) end of September prediction market resolve?
This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET. Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. The settlement source listed for this market is livebench.ai.