Tech AI

Which company has the best AI model on LiveBench (Coding) end of September?

Sort by
Anthropic
$4.63K Vol.
54.5% 1%
OpenAI
$9.72K Vol.
31.5% 1%
Meta
$1.41K Vol.
4.4%
SpaceXAI
$1.17K Vol.
0.7%
Nvidia
$1.06K Vol.
0.2%
16 more outcomes Listed by current odds, highest first

Odds summary

Anthropic currently leads the Which company has the best AI model on LiveBench (Coding) end of September prediction market at 54.5% reported probability on Polymarket. The figures below combine live odds, liquidity, volume, and open interest so readers can compare the market signal before reading the full analysis.

Volume$29.89K Liquidity$10.37K Open Interest$6.62K Last updated8 mins ago

Odds, liquidity, volume, and open interest are sourced from Polymarket and last synced at Aug 8, 2026 1:42 am.

CryptoSlate Market Analysis

Anthropic’s Lead Prices Persistence While OpenAI Retains Release Optionality

Claude Fable 5’s benchmark lead gives Anthropic the stronger base case, while GPT-5.6 Sol’s recent and incomplete rollout preserves a credible path for OpenAI. The key question is whether September’s winner comes from steady iteration or a late model update.

Humanoid AI operator evaluating three illuminated computing towers in a futuristic benchmark arena.

Anthropic’s 60.5% probability is best understood as a wager on benchmark persistence: Claude Fable 5 already leads LiveBench Coding, and the remaining window may favor the incumbent unless OpenAI converts a recent release into a higher score. OpenAI’s 35.5% share prices a specific alternative—GPT-5.6 Sol improves through deployment, iteration, or another eligible release before the September 30 snapshot.

Fable 5’s existing lead gives Anthropic the cleaner path

The latest public LiveBench leaderboard places Anthropic’s Claude Fable 5 ahead of OpenAI’s GPT-5.6 Sol in Coding. That sourced fact explains most of the hierarchy: Anthropic can prevail if the ordering survives, while OpenAI needs the leaderboard to change.

Anthropic also describes Claude Fable 5, restored and rolled out globally on July 1, as its most capable model for ambitious coding projects. A broadly deployed model gives Anthropic an established benchmark candidate and time to refine the surrounding system. The market inference is that Fable 5’s score has enough durability to withstand near-term competition. The 60.5% price still leaves substantial room for displacement, suggesting the current lead alone is insufficient to settle expectations months early.

OpenAI’s probability rests on a release still entering circulation

OpenAI launched GPT-5.6 on July 9 and calls GPT-5.6 Sol its best coding model yet, citing state-of-the-art performance across coding and related tasks. Its July 9 ChatGPT release notes say access is rolling out gradually to eligible plans and may not yet reach every account.

That rollout detail matters because the current leaderboard may capture an early point in Sol’s deployment cycle. A wider release could coincide with configuration changes, improved tool use, or a refreshed submission, although none of those outcomes is confirmed by the supplied evidence. OpenAI’s 35.5% probability therefore represents meaningful release optionality alongside a present benchmark deficit. Evidence that Sol’s score rises after broader availability would strengthen this interpretation. A stable score despite completed rollout would weaken it.

The pricing assumes a two-lab contest through September

Every listed company outside Anthropic and OpenAI sits at 0.7% or below, with most at 0.1%. The implied story is that Meta, Google, Alibaba, DeepSeek, Amazon, Nvidia, and other listed developers are unlikely to introduce an eligible model capable of taking first place by the check date. This is a strong hidden assumption because the contract rewards the single highest Coding score at one specified moment, allowing a late release to matter even without a long public track record.

The available record supplies no comparable upcoming release evidence for those companies, so their low probabilities should be read as market inference, not proof that they lack competitive internal models. A dated model announcement, an official benchmark submission, or a sudden LiveBench entry near the leaders would challenge the two-lab framing immediately.

Resolution mechanics make timing and model ownership decisive

The contract resolves to the company owning the model with the highest LiveBench Coding score when checked September 30, 2026, at 12:00 p.m. ET. This creates a snapshot contest. Average performance across the preceding months, adoption, price, and developer preference do not determine settlement under the stated rule.

A model update arriving shortly before the check can therefore outweigh months of leaderboard stability. The ownership wording also makes the identity attached to the leading model material. Any renamed model, partnership, acquisition, or jointly developed system would require applying the published resolution language to the leaderboard entry.

Score stability is the main counter-signal to an OpenAI comeback

The strongest counterargument to OpenAI’s path is simple: its newest coding model is already public, yet Fable 5 remains ahead in the supplied leaderboard record. If GPT-5.6 Sol completes rollout without narrowing the gap, Anthropic’s persistence thesis gains support. Anthropic could also release an updated Fable variant, extending the target OpenAI must clear.

Concrete repricing catalysts include a LiveBench score refresh, an official model revision from either lab, confirmation that GPT-5.6 Sol’s rollout is complete, or a new entrant reaching the top tier. The market’s $28,080 volume, $21,420 liquidity, and $6,990 open interest support a visible hierarchy, though they provide limited evidence that distant release schedules have been fully incorporated. Until a score or release changes, the current ordering favors the company already leading the settlement benchmark.

Sources

What could move the odds?

Informational summary of factors that may affect the reported prediction-market probabilities.

Market-implied thesis

At 54.5%, the market treats Anthropic as more likely than any named rival to own the model atop LiveBench Coding at the September check, not as a settled lead.

The current LiveBench release reportedly favors Anthropic, while its Fable 5 and Opus 4.8 materials emphasize coding work. The price still leaves substantial combined probability for challengers.

Mixed signal 66% CatalystLiveBench leaderboard check on September 30, 2026 RiskLeaderboard leadership can change before the check.

What could reprice it

The September 30 LiveBench leaderboard check is the decisive dated event: a newly listed model that leads Coding by then would directly determine the winning company.

Polymarket resolves on the company owning the highest Coding-scoring model when LiveBench.ai is checked at 12:00 PM ET, making benchmark inclusion and rank changes more material than launch claims alone.

Strong signal 82% CatalystSeptember 30, 2026, 12:00 PM ET leaderboard check RiskNew models may not be reflected on the leaderboard.

Where the market may be weak

The displayed $12.64K liquidity does not establish broad, durable participation, so quoted probabilities may move materially without representing a deep consensus on benchmark leadership.

The page reports $29.89K volume but no trader count. Cumulative turnover is not the same as executable depth, and neither figure demonstrates that informed benchmark observers set the prices.

Thin signal 40% CatalystAdditional participation or leaderboard updates RiskLimited depth can amplify price moves.

Counter-signal

Anthropic’s lead could fail if OpenAI’s coding models outperform on LiveBench: OpenAI calls GPT-5.5 and GPT-5.3-Codex its strongest agentic coding models to date.

OpenAI is the clearest priced challenger at 30%. Google also announced improved coding performance for Gemini 3.6 Flash, showing that the relevant model race remains active rather than fixed.

Mixed signal 60% CatalystA challenger model reaching the LiveBench Coding leaderboard RiskCompany launch claims are not LiveBench scores.

Market details

Resolution criteria
This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET.
Platform
Category
Tech AI
Close date
September 30, 2026, 11:59 PM UTC
Settlement source
livebench.ai
Market rules summary
Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. View full rules

Frequently asked questions

What are the current Which company has the best AI model on LiveBench (Coding) end of September odds?

Polymarket reports Which company has the best AI model on LiveBench (Coding) end of September odds with Anthropic at 54.5%, OpenAI at 31.5%, Meta at 4.4%, and SpaceXAI at 0.7%. These probabilities are market-implied and can change as liquidity and trading activity update. The latest market snapshot includes $29.89K volume, $10.37K liquidity, and $6.62K open interest. CryptoSlate last synced this market data at Aug 8, 2026, 00:42 UTC.

What could move the Which company has the best AI model on LiveBench (Coding) end of September prediction market odds?

At 54.5%, the market treats Anthropic as more likely than any named rival to own the model atop LiveBench Coding at the September check, not as a settled lead. The current LiveBench release reportedly favors Anthropic, while its Fable 5 and Opus 4.8 materials emphasize coding work. The price still leaves substantial combined probability for challengers. Catalysts to watch include LiveBench leaderboard check on September 30, 2026, September 30, 2026, 12:00 PM ET leaderboard check, and Additional participation or leaderboard updates.

How does the Which company has the best AI model on LiveBench (Coding) end of September prediction market resolve?

This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET. Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. The settlement source listed for this market is livebench.ai.

Follow the signal

Curated intelligence, delivered your way.

Never miss a market-moving update.

  • Daily briefingTop stories & analysis
  • Market movesKey charts & data
  • Policy updatesWhat to watch
  • Weekly deep diveLong-form insights

Subscribe to the CryptoSlate newsletter through Substack.

Join 100,000+ readers Unsubscribe anytime