Which company has the best AI model on LiveBench (Coding) end of September?
Anthropic is the current LiveBench Coding frontrunner, and Claude Fable 5 plus Opus 4.8 are both positioned around serious coding and agentic work. If those models keep topping the leaderboard through the Sept.
A stronger OpenAI or Google coding release before settlement could overtake Claude on LiveBench and push Anthropic out of first.
AI-Assisted. May contain errors.
OpenAI has a direct catalyst in GPT-5.5 and GPT-5.3-Codex, both described as its strongest or most capable agentic coding models to date. If one of those releases translates into a LiveBench Coding lead before Sept.
If OpenAI’s newer coding models fail to beat Anthropic’s score, the market’s main challenger thesis weakens and first place stays elsewhere.
AI-Assisted. May contain errors.
Meta would need a late coding-model jump from its frontier stack to beat Anthropic and OpenAI on LiveBench Coding by the Sept. 30 check.
Without a clearly superior new coding model, Meta is likely to remain behind the current leaders and miss the top LiveBench score.
AI-Assisted. May contain errors.
16 more outcomes Listed by current odds, highest first
Odds summary
Anthropic currently leads the Which company has the best AI model on LiveBench (Coding) end of September prediction market at 54.5% reported probability on Polymarket. The figures below combine live odds, liquidity, volume, and open interest so readers can compare the market signal before reading the full analysis.
Odds, liquidity, volume, and open interest are sourced from Polymarket and last synced at Aug 8, 2026 1:42 am.
Anthropic’s Lead Prices Persistence While OpenAI Retains Release Optionality
Claude Fable 5’s benchmark lead gives Anthropic the stronger base case, while GPT-5.6 Sol’s recent and incomplete rollout preserves a credible path for OpenAI. The key question is whether September’s winner comes from steady iteration or a late model update.

Anthropic’s 60.5% probability is best understood as a wager on benchmark persistence: Claude Fable 5 already leads LiveBench Coding, and the remaining window may favor the incumbent unless OpenAI converts a recent release into a higher score. OpenAI’s 35.5% share prices a specific alternative—GPT-5.6 Sol improves through deployment, iteration, or another eligible release before the September 30 snapshot.
Fable 5’s existing lead gives Anthropic the cleaner path
The latest public LiveBench leaderboard places Anthropic’s Claude Fable 5 ahead of OpenAI’s GPT-5.6 Sol in Coding. That sourced fact explains most of the hierarchy: Anthropic can prevail if the ordering survives, while OpenAI needs the leaderboard to change.
Anthropic also describes Claude Fable 5, restored and rolled out globally on July 1, as its most capable model for ambitious coding projects. A broadly deployed model gives Anthropic an established benchmark candidate and time to refine the surrounding system. The market inference is that Fable 5’s score has enough durability to withstand near-term competition. The 60.5% price still leaves substantial room for displacement, suggesting the current lead alone is insufficient to settle expectations months early.
OpenAI’s probability rests on a release still entering circulation
OpenAI launched GPT-5.6 on July 9 and calls GPT-5.6 Sol its best coding model yet, citing state-of-the-art performance across coding and related tasks. Its July 9 ChatGPT release notes say access is rolling out gradually to eligible plans and may not yet reach every account.
That rollout detail matters because the current leaderboard may capture an early point in Sol’s deployment cycle. A wider release could coincide with configuration changes, improved tool use, or a refreshed submission, although none of those outcomes is confirmed by the supplied evidence. OpenAI’s 35.5% probability therefore represents meaningful release optionality alongside a present benchmark deficit. Evidence that Sol’s score rises after broader availability would strengthen this interpretation. A stable score despite completed rollout would weaken it.
The pricing assumes a two-lab contest through September
Every listed company outside Anthropic and OpenAI sits at 0.7% or below, with most at 0.1%. The implied story is that Meta, Google, Alibaba, DeepSeek, Amazon, Nvidia, and other listed developers are unlikely to introduce an eligible model capable of taking first place by the check date. This is a strong hidden assumption because the contract rewards the single highest Coding score at one specified moment, allowing a late release to matter even without a long public track record.
The available record supplies no comparable upcoming release evidence for those companies, so their low probabilities should be read as market inference, not proof that they lack competitive internal models. A dated model announcement, an official benchmark submission, or a sudden LiveBench entry near the leaders would challenge the two-lab framing immediately.
Resolution mechanics make timing and model ownership decisive
The contract resolves to the company owning the model with the highest LiveBench Coding score when checked September 30, 2026, at 12:00 p.m. ET. This creates a snapshot contest. Average performance across the preceding months, adoption, price, and developer preference do not determine settlement under the stated rule.
A model update arriving shortly before the check can therefore outweigh months of leaderboard stability. The ownership wording also makes the identity attached to the leading model material. Any renamed model, partnership, acquisition, or jointly developed system would require applying the published resolution language to the leaderboard entry.
Score stability is the main counter-signal to an OpenAI comeback
The strongest counterargument to OpenAI’s path is simple: its newest coding model is already public, yet Fable 5 remains ahead in the supplied leaderboard record. If GPT-5.6 Sol completes rollout without narrowing the gap, Anthropic’s persistence thesis gains support. Anthropic could also release an updated Fable variant, extending the target OpenAI must clear.
Concrete repricing catalysts include a LiveBench score refresh, an official model revision from either lab, confirmation that GPT-5.6 Sol’s rollout is complete, or a new entrant reaching the top tier. The market’s $28,080 volume, $21,420 liquidity, and $6,990 open interest support a visible hierarchy, though they provide limited evidence that distant release schedules have been fully incorporated. Until a score or release changes, the current ordering favors the company already leading the settlement benchmark.
Sources
What could move the odds?
Informational summary of factors that may affect the reported prediction-market probabilities.
Market-implied thesis
At 54.5%, the market treats Anthropic as more likely than any named rival to own the model atop LiveBench Coding at the September check, not as a settled lead.
The current LiveBench release reportedly favors Anthropic, while its Fable 5 and Opus 4.8 materials emphasize coding work. The price still leaves substantial combined probability for challengers.
What could reprice it
The September 30 LiveBench leaderboard check is the decisive dated event: a newly listed model that leads Coding by then would directly determine the winning company.
Polymarket resolves on the company owning the highest Coding-scoring model when LiveBench.ai is checked at 12:00 PM ET, making benchmark inclusion and rank changes more material than launch claims alone.
Where the market may be weak
The displayed $12.64K liquidity does not establish broad, durable participation, so quoted probabilities may move materially without representing a deep consensus on benchmark leadership.
The page reports $29.89K volume but no trader count. Cumulative turnover is not the same as executable depth, and neither figure demonstrates that informed benchmark observers set the prices.
Counter-signal
Anthropic’s lead could fail if OpenAI’s coding models outperform on LiveBench: OpenAI calls GPT-5.5 and GPT-5.3-Codex its strongest agentic coding models to date.
OpenAI is the clearest priced challenger at 30%. Google also announced improved coding performance for Gemini 3.6 Flash, showing that the relevant model race remains active rather than fixed.
Market details
- Resolution criteria
- This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET.
- Category
- Tech › AI
- Close date
- September 30, 2026, 11:59 PM UTC
- Settlement source
- livebench.ai
- Market rules summary
- Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. View full rules
Frequently asked questions
What are the current Which company has the best AI model on LiveBench (Coding) end of September odds?
Polymarket reports Which company has the best AI model on LiveBench (Coding) end of September odds with Anthropic at 54.5%, OpenAI at 31.5%, Meta at 4.4%, and SpaceXAI at 0.7%. These probabilities are market-implied and can change as liquidity and trading activity update. The latest market snapshot includes $29.89K volume, $10.37K liquidity, and $6.62K open interest. CryptoSlate last synced this market data at Aug 8, 2026, 00:42 UTC.
What could move the Which company has the best AI model on LiveBench (Coding) end of September prediction market odds?
At 54.5%, the market treats Anthropic as more likely than any named rival to own the model atop LiveBench Coding at the September check, not as a settled lead. The current LiveBench release reportedly favors Anthropic, while its Fable 5 and Opus 4.8 materials emphasize coding work. The price still leaves substantial combined probability for challengers. Catalysts to watch include LiveBench leaderboard check on September 30, 2026, September 30, 2026, 12:00 PM ET leaderboard check, and Additional participation or leaderboard updates.
How does the Which company has the best AI model on LiveBench (Coding) end of September prediction market resolve?
This market will resolve according to the company which owns the model with the highest Coding score on LiveBench.ai when the leaderboard is checked on September 30, 2026, at 12:00 PM ET. Multi-outcome Polymarket event. Each listed option is represented by its Yes price on the underlying market. The settlement source listed for this market is livebench.ai.