Tech AI

Which company has the best AI model on LiveBench (Coding) end of September?

Sort by
Anthropic
$4.63K Vol.
54.5% 1%
OpenAI
$9.66K Vol.
30.5% 5.5%
Meta
$1.41K Vol.
4.9% 0.4%
SpaceXAI
$1.17K Vol.
0.7%
Nvidia
$1.06K Vol.
0.2%
16 more outcomes Listed by current odds, highest first

Odds summary

Anthropic currently leads the Which company has the best AI model on LiveBench (Coding) end of September prediction market at 54.5% reported probability on Polymarket. The figures below combine live odds, liquidity, volume, and open interest so readers can compare the market signal before reading the full analysis.

Volume$29.83K Liquidity$12.46K Open Interest$6.64K Last updated13 mins ago

Odds, liquidity, volume, and open interest are sourced from Polymarket and last synced at Aug 7, 2026 11:47 pm.

CryptoSlate Market Analysis

Anthropic’s Lead Prices Persistence While OpenAI Retains Release Optionality

Claude Fable 5’s benchmark lead gives Anthropic the stronger base case, while GPT-5.6 Sol’s recent and incomplete rollout preserves a credible path for OpenAI. The key question is whether September’s winner comes from steady iteration or a late model update.

Humanoid AI operator evaluating three illuminated computing towers in a futuristic benchmark arena.

Anthropic’s 60.5% probability is best understood as a wager on benchmark persistence: Claude Fable 5 already leads LiveBench Coding, and the remaining window may favor the incumbent unless OpenAI converts a recent release into a higher score. OpenAI’s 35.5% share prices a specific alternative—GPT-5.6 Sol improves through deployment, iteration, or another eligible release before the September 30 snapshot.

Fable 5’s existing lead gives Anthropic the cleaner path

The latest public LiveBench leaderboard places Anthropic’s Claude Fable 5 ahead of OpenAI’s GPT-5.6 Sol in Coding. That sourced fact explains most of the hierarchy: Anthropic can prevail if the ordering survives, while OpenAI needs the leaderboard to change.

Anthropic also describes Claude Fable 5, restored and rolled out globally on July 1, as its most capable model for ambitious coding projects. A broadly deployed model gives Anthropic an established benchmark candidate and time to refine the surrounding system. The market inference is that Fable 5’s score has enough durability to withstand near-term competition. The 60.5% price still leaves substantial room for displacement, suggesting the current lead alone is insufficient to settle expectations months early.

OpenAI’s probability rests on a release still entering circulation

OpenAI launched GPT-5.6 on July 9 and calls GPT-5.6 Sol its best coding model yet, citing state-of-the-art performance across coding and related tasks. Its July 9 ChatGPT release notes say access is rolling out gradually to eligible plans and may not yet reach every account.

That rollout detail matters because the current leaderboard may capture an early point in Sol’s deployment cycle. A wider release could coincide with configuration changes, improved tool use, or a refreshed submission, although none of those outcomes is confirmed by the supplied evidence. OpenAI’s 35.5% probability therefore represents meaningful release optionality alongside a present benchmark deficit. Evidence that Sol’s score rises after broader availability would strengthen this interpretation. A stable score despite completed rollout would weaken it.

The pricing assumes a two-lab contest through September

Every listed company outside Anthropic and OpenAI sits at 0.7% or below, with most at 0.1%. The implied story is that Meta, Google, Alibaba, DeepSeek, Amazon, Nvidia, and other listed developers are unlikely to introduce an eligible model capable of taking first place by the check date. This is a strong hidden assumption because the contract rewards the single highest Coding score at one specified moment, allowing a late release to matter even without a long public track record.

The available record supplies no comparable upcoming release evidence for those companies, so their low probabilities should be read as market inference, not proof that they lack competitive internal models. A dated model announcement, an official benchmark submission, or a sudden LiveBench entry near the leaders would challenge the two-lab framing immediately.

Resolution mechanics make timing and model ownership decisive

The contract resolves to the company owning the model with the highest LiveBench Coding score when checked September 30, 2026, at 12:00 p.m. ET. This creates a snapshot contest. Average performance across the preceding months, adoption, price, and developer preference do not determine settlement under the stated rule.

A model update arriving shortly before the check can therefore outweigh months of leaderboard stability. The ownership wording also makes the identity attached to the leading model material. Any renamed model, partnership, acquisition, or jointly developed system would require applying the published resolution language to the leaderboard entry.

Score stability is the main counter-signal to an OpenAI comeback

The strongest counterargument to OpenAI’s path is simple: its newest coding model is already public, yet Fable 5 remains ahead in the supplied leaderboard record. If GPT-5.6 Sol completes rollout without narrowing the gap, Anthropic’s persistence thesis gains support. Anthropic could also release an updated Fable variant, extending the target OpenAI must clear.

Concrete repricing catalysts include a LiveBench score refresh, an official model revision from either lab, confirmation that GPT-5.6 Sol’s rollout is complete, or a new entrant reaching the top tier. The market’s $28,080 volume, $21,420 liquidity, and $6,990 open interest support a visible hierarchy, though they provide limited evidence that distant release schedules have been fully incorporated. Until a score or release changes, the current ordering favors the company already leading the settlement benchmark.

Sources

What could move the odds?

Informational summary of factors that may affect the reported prediction-market probabilities.

Market-implied thesis

Market pricing implies Anthropic is likelier than any single rival to own the model leading LiveBench Coding at the specified check.

This is a relative winner claim, not a claim that Anthropic will retain a permanent coding advantage. The research context identifies Anthropic as the current leader in the latest release.

Mixed signal 65% CatalystModel releases or a leaderboard refresh before September 30 RiskLeadership can change before the observation time

What could reprice it

A stronger coding-model release or a LiveBench leaderboard refresh before the September check is the clearest route to material repricing.

OpenAI describes GPT-5.6 Sol as stronger in coding, while its prior GPT-5.3-Codex release shows continuing coding-model iteration. Neither source supplies a dated release or benchmark-refresh schedule.

Mixed signal 58% CatalystBroader GPT-5.6 Sol availability or a LiveBench refresh RiskNo specific release or refresh date is provided

Where the market may be weak

The deciding leaderboard snapshot precedes the listed close time, leaving unclear whether trading can continue after the observation that fixes settlement.

Rules set the leaderboard check for September 30 at 12:00 PM ET, while the listed close is September 30 at 11:59 PM UTC. That timing gap may complicate interpretation of post-snapshot price discovery.

Rules risk 55% CatalystClarification of trading and resolution timing RiskSnapshot and close-time sequencing is not fully explained

Counter-signal

OpenAI could displace Anthropic if GPT-5.6 Sol is broadly exposed and records the highest LiveBench Coding score by the check.

OpenAI says GPT-5.6 Sol has stronger coding capabilities, and its February 5 GPT-5.3-Codex launch was described as its most capable agentic coding model. The supplied sources do not establish a LiveBench score for either model.

Mixed signal 61% CatalystOpenAI model exposure and benchmark inclusion RiskCapability claims may not translate into leaderboard leadership

Market details

Resolution criteria
This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET.
Platform
Category
Tech AI
Close date
September 30, 2026, 11:59 PM UTC
Settlement source
livebench.ai
Market rules summary
Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. View full rules

Frequently asked questions

What are the current Which company has the best AI model on LiveBench (Coding) end of September odds?

Polymarket reports Which company has the best AI model on LiveBench (Coding) end of September odds with Anthropic at 54.5%, OpenAI at 30.5%, Meta at 4.9%, and SpaceXAI at 0.7%. These probabilities are market-implied and can change as liquidity and trading activity update. The latest market snapshot includes $29.83K volume, $12.46K liquidity, and $6.64K open interest. CryptoSlate last synced this market data at Aug 7, 2026, 22:47 UTC.

What could move the Which company has the best AI model on LiveBench (Coding) end of September prediction market odds?

Market pricing implies Anthropic is likelier than any single rival to own the model leading LiveBench Coding at the specified check. This is a relative winner claim, not a claim that Anthropic will retain a permanent coding advantage. The research context identifies Anthropic as the current leader in the latest release. Catalysts to watch include Model releases or a leaderboard refresh before September 30, Broader GPT-5.6 Sol availability or a LiveBench refresh, and Clarification of trading and resolution timing.

How does the Which company has the best AI model on LiveBench (Coding) end of September prediction market resolve?

This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET. Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. The settlement source listed for this market is livebench.ai.

Follow the signal

Curated intelligence, delivered your way.

Never miss a market-moving update.

  • Daily briefingTop stories & analysis
  • Market movesKey charts & data
  • Policy updatesWhat to watch
  • Weekly deep diveLong-form insights

Subscribe to the CryptoSlate newsletter through Substack.

Join 100,000+ readers Unsubscribe anytime