As the final quarter of 2026 approaches, the global battle for artificial intelligence supremacy is no longer waged solely in academic benchmarks or corporate keynotes. Today, financial prediction platforms like Kalshi provide a high-stakes, real-time snapshot of market sentiment. Traders are putting millions on the line to forecast which laboratory will hold the top position on the LMArena text leaderboard when the clock strikes midnight on December 31, 2026.
Despite aggressive moves from rivals and unprecedented regulatory oversight throughout the summer, Anthropic’s Claude maintains a commanding lead, holding nearly two-thirds of the market's implied probability. However, the recent deployment of OpenAI’s GPT-6 Astra has sparked renewed volatility across both general and domain-specific prediction boards.
Data from Kalshi reveals a stark division between the top two contenders and the rest of the industry. The contract pays out based on which company controls the number-one ranked Large Language Model (LLM) on LMArena’s style-control-removed benchmark at year's end.
- Claude (Anthropic) : 62¢ (~62%)
- ChatGPT (OpenAI) : 25¢ (~25%)
- Gemini (Google) : 6¢ (~6%)
- Grok (xAI) : 5¢ (~5%)
- Muse Spark (Meta) : 1¢ (~1%)
Anthropic’s dominance is grounded in tangible leaderboard stats. Its flagship claude-fable-5 sits at the top of the LMArena rankings with an Arena score of 1507. In fact, Anthropic models currently occupy seven of the top nine spots on the public leaderboard.
While Google recently launched Gemini 3.8—earning praise for its extended context recall and raw execution speeds—traders remain skeptical that incremental updates will close the gap in time, pricing Google at just 6¢.
On September 3, 2026, OpenAI attempted to shift the momentum by deploying GPT-6 Astra. Rollout began with limited enterprise availability before expanding to Plus and Pro tier accounts. OpenAI President Greg Brockman framed the milestone assertively, remarking that "it's not unreasonable to feel that we are now in the AGI era," while official documentation designated Astra as reaching a "Critical" cybersecurity threshold under their internal Preparedness Framework.
The launch triggered heavy volume, with over 274,000 contracts changing hands on Kalshi within days. Interestingly, the flagship "Best AI" year-end contract absorbed the release with a modest four-cent gain for OpenAI, bringing ChatGPT from 21¢ to 25¢.
However, a dramatic split emerged on Kalshi’s specialized Coding Model Market (which settles on LiveBench’s Coding Average):
OpenAI’s Coding Contract: Surged by 11 cents in bid pricing, jumping from 17¢/18¢ up to 28¢/33¢ post-launch.
Anthropic’s Coding Contract: Slipped from 69¢/73¢ down to 61¢/66¢ over the same weekend window.
This divergence illustrates that while traders view GPT-6 Astra as a formidable leap forward in programmatic execution, they are withholding judgment on whether it can unseat Claude on broader language and reasoning leaderboards before the December 31 deadline.
Anthropic’s current favorite status is particularly impressive given the regulatory turbulence it faced earlier in the summer. In mid-2026, US federal authorities briefly restricted foreign deployments of Claude Fable 5 and Mythos 5 due to systemic cybersecurity concerns regarding autonomous exploitation risks.
The setback proved temporary. Following a comprehensive audit by the US Center for AI Standards and Innovation (CAISI), Anthropic secured full clearance. Market confidence surged again after Anthropic published a detailed safety assessment demonstrating how their internal safeguards intercepted attempts to adapt Claude models for chemical and biological threat generation.
This recovery highlights a crucial shift in the 2026 AI ecosystem: institutional investors and prediction markets now equate aggressive safety engineering directly with commercial resilience and long-term valuation.
As the market enters its final quarter, a compelling tension exists between corporate reach and benchmark performance. OpenAI continues to dominate consumer mindshare, boasting over 1 billion monthly active users and an annual ad revenue run rate topping $1 billion.
Yet, prediction markets care strictly about benchmark victory on December 31. With OpenAI’s current scored entry sitting at 17th place (gpt-5.6-sol-xhigh), the 25¢ price tag on ChatGPT represents a forward-looking bet that GPT-6 Astra will score exceptionally high once fully evaluated on LMArena.
Whether OpenAI can complete that climb—or if Anthropic has another surprise release waiting in the wings—will determine who takes the crown in 2026.
Sed at tellus, pharetra lacus, aenean risus non nisl ultricies commodo diam aliquet arcu enim eu leo porttitor habitasse adipiscing porttitor varius ultricies facilisis viverra lacus neque.



