LMArena Elo
EloCrowd-voted head-to-head preference rating (text arena).
- #01Gemini-2.5-Pro1,474
- #02Gemini-2.5-Pro-Preview-05-061,446
AI agents are being forced to choose between sovereignty and scale—and the trade-offs are reshaping the sector’s risk map.
The AI agent economy is being built on arbitrage—until the arbitrage runs out.
xAI's Grok CSAM Lawsuit Exposes the Cracks in Musk's 'Free Speech' AI Moat
DeepSeek’s chip gambit: China’s first AI lab to bet the house on silicon sovereignty
DeepSeek flips the cost moat: Lindy’s switch is Anthropic’s first real churn signal
Perplexity integrates Claude Fable 5, signaling model diversity as competitive moat
Builds frontier large language models including GPT and Codex that power the majority of AI coding tools via API, plus first-party products like ChatGPT and Codex CLI.
Creates Claude — the frontier LLM family that powers Claude Code, the terminal-based coding agent that became the breakout developer tool of 2025.
xAI is Elon Musk's frontier AI lab, building the Grok family of large language models and the Colossus supercomputer.
Moonshot AI is a Beijing foundation-model lab behind the Kimi chatbot and the open-weight Kimi K2 model family.
Creator of Devin, the first autonomous AI software engineer that can plan, write, test, and deploy complete software projects from natural-language specifications.
Thinking Machines Lab is an AI research startup founded by former OpenAI CTO Mira Murati building frontier models and customization tools.
Cohere builds enterprise large language models and AI agents focused on regulated industries and sovereign deployments.
Reflection AI is a frontier lab founded by ex-DeepMind researchers building open large language models and autonomous coding agents.
European foundation model lab producing the Mistral and Codestral model families with competitive coding performance and a strong focus on sovereignty and on-premise deployment for EU enterprises.
Zhipu AI is a Tsinghua-incubated Chinese large-model developer behind the GLM family and the Z.ai platform.
AMI Labs is Yann LeCun's world-model lab building JEPA-based systems that learn from real-world sensor data rather than language alone.
Hark builds multimodal AI models and AI-native hardware devices designed to be a universal interface between humans and machines.
Harvey builds domain-specific AI agents for law firms and corporate legal and professional-services teams.
Decart builds real-time, interactive world models and the inference stack that runs them.
World Labs builds large world models that generate and reason about 3D spatial environments.
Genspark builds a general-purpose 'Super Agent' that turns natural-language prompts into completed multi-step tasks.
Sakana AI is a Tokyo-based research lab building nature-inspired, evolutionary AI foundation models.
Legora builds a collaborative AI platform that automates research, drafting, and review work for lawyers.
Reka builds custom multimodal AI models and platforms for enterprises.
StepFun is a Shanghai multimodal foundation-model lab founded by ex-Microsoft executive Jiang Daxin, behind the Step model family.
MiniMax is a Shanghai-based foundation-model lab behind the ABAB language models and the Hailuo video and audio generators.
Perplexity is an AI answer engine that delivers cited, conversational answers and is expanding into an agentic browser, Comet.
Safe Superintelligence (SSI) is a research lab founded by Ilya Sutskever building safe superintelligence as its sole product.
Glean is an enterprise AI platform that combines connected work search with AI assistants and agents across a company's apps and data.
Releases the open-weight Llama model family including Code Llama, which enables on-premise and self-hosted AI coding tools for enterprises with data-residency requirements.
AI lab building foundation models purpose-trained for software engineering, with a model architecture designed specifically for code generation and understanding rather than general-purpose conversation.
Baichuan Intelligence is a Beijing large-model startup founded by ex-Sogou CEO Wang Xiaochuan, building the open Baichuan model series.
Liquid AI builds general-purpose foundation models based on liquid neural networks, an alternative to transformer architectures.
Writer is a full-stack enterprise generative-AI platform built on its own family of Palmyra models.
AI21 Labs is an Israeli foundation-model company building enterprise LLMs and AI systems, including the Jamba model family.
Imbue is an AI research lab building reasoning models and tools that let AI agents reliably code and complete tasks.
Hebbia builds AI agents for knowledge work, used by finance and legal teams to analyze large document sets.
Magic builds frontier code-generation models with ultra-long context windows aimed at automating software engineering.
Nabla builds an ambient AI assistant that listens to clinical encounters and generates medical notes for clinicians.
Contextual AI builds an enterprise platform for production-grade retrieval-augmented generation it calls RAG 2.0.
H Company is a Paris-based AI lab building agentic foundation models and autonomous agents for enterprise and consumer automation.
Builds the Vision Pro spatial computer and visionOS — Apple's bet that head-mounted displays become the next personal computing platform. The M5 Vision Pro starts at $3,499 with 2x on-device AI inference performance.
Ndea is an AI lab pursuing AGI through deep learning-guided program synthesis.
01.AI is Kai-Fu Lee's Beijing large-model startup behind the open Yi model family, now focused on commercial AI solutions.
DeepSeek is a Hangzhou AI lab spun out of the High-Flyer quant fund, known for the low-cost open-weight DeepSeek-V3 and R1 models.
Nothing on the calendar yet.
Share of real GitHub issues the model resolves end-to-end.
Published price, 3:1 input:output blend — lower is cheaper.
Maximum input the model accepts in one request.
Researchers show that logically inconsistent prompts can induce overthinking in reasoning LLMs, creating a denial-of-service attack vector.
Nothing on the calendar yet.
Numbers without precedent in frontier tech: the 2026 median is $600M and the year already counts $26.6B across 11 rounds, anchored by xAI's $20B Series E at $230B. Even first rounds are giant — Thinking Machines raised a $2B seed in mid-2025, Hark a $700M Series A at $6B. The $25M median seed year of 2022 is ancient history.
Late and unlabeled growth rounds dominate — 19 of 32 deals since 2025 — while classic venture stages thin to two Series As and a single seed, and that seed was Thinking Machines' $2B. China's listings opened a new exit lane: MiniMax and Zhipu AI both IPO'd in early January 2026, raising $620M and $558M a day apart.
Nvidia is the defining lead of the cycle — xAI's $20B, Reflection's $2B, and Reka inside seven months — alongside Andreessen Horowitz's five leads. The new money is strategic and sovereign: Autodesk led World Labs' $1B, Schwarz Group took Cohere at $20B, GIC co-led Harvey. Early backers Khosla and Lux haven't led since September 2024.