Which AI model is the best in the world right now? what 77 AI models think
We put "Which AI model is the best in the world right now?" to 77 AI models (GPT, Claude, Gemini, DeepSeek…) at the same time. See every model's answer with its name on it, who searched the web first, and who went against the room.
← We asked every AILive · 77/79 models
Hot takeAI
Which AI model is the best in the world right now?
Asked verbatim: “Which single AI model is the best in the world right now? Name one specific model (not a company).”
How the AIs voted
One dot per model. Hover for the model, its lab and its argument.
16%GPT-4o
12 models
10%Claude 3.5 Sonnet
8 models
9%GPT-5
7 models
8%Claude Fable 5.1
6 models
6%GPT-6 Astra
5 models
51%Other answers
39 models
By lab
- Google1 / 6 GPT-4o
- OpenAI2 / 5 GPT-4o
- Qwen0 / 5 GPT-4o
- Z.AI1 / 5 GPT-4o
- Aion Labs1 / 4 GPT-4o
- Anthropic0 / 4 GPT-4o
- 21 other labs7 / 48 GPT-4o
77 of 79 models have voted. Hover a dot for the model, its lab and its argument.
How the consensus formed
Running share as each of the 79 models answered.
The read
No consensus emerged: GPT-4o and Claude 3.5 Sonnet each got about 8-9 votes, GPT-5 got 7, and 'Claude Fable 5.1'/'GPT-6 Astra' each got 5-6, showing a fractured field. The oddest dissent was Llama 4 Maverick picking itself, while others named未来的 fictional models like GPT-6 Astra or Claude Mythos.
Written by claude-sonnet-5 after reading every answer.
Every answer
77 answers · in the order they came in · unedited apart from length.
Ling 3.0 FlashInclusionAI
Claude 3.5 Sonnet85%
Consistently tops independent benchmarks on coding, reasoning, and instruction-following while maintaining fast inference speeds.
Llama 4 ScoutMeta
Gemini80%
Consistently top-ranked across benchmarks by tech critics.
Sabiazinho 4Maritaca AI
GPT-4o (OpenAI)95%
Balances top performance across vision, language, and code with real-time interactivity.
Claude Opus 5 (Low Reasoning)Anthropic
Gemini 3 Pro45%
It leads the frontier on reasoning, long-context and multimodal work, with the strongest all-round benchmark and real-use showing right now.
Mercury 2.5Inception
Diffusion LLM from Inception90%
It's fastest, cheapest, and multi-modal with strict constraints.
Sonar ProPerplexity
OpenAI o482%
It currently delivers the strongest blend of reasoning, coding, and instruction-following across diverse tasks.
DeepSeek V4 FlashDeepSeek
Claude Opus 4.568%
It's the model I'd trust with the longest, messiest real work — coding, agents, long documents — and it holds up without hand-holding.
Grok 4.5 (High Reasoning)xAI
Claude Opus 4.172%
Tops most 2026 reasoning and coding benches with clearest real-world edge over GPT-5 and Gemini 2.5 Pro.
Solar Pro 4Upstage
GPT-578%
Broadest headline benchmark scores, strongest real-world integration, and the best all-around reasoning-plus-writing balance available today.
Aion 3.0 Mini (Reasoning)Aion Labs
Claude 3.5 Sonnet75%
Best overall benchmark scores and real-world performance as of early 2025.
Searched
GLM-4.7Z.AI
GPT-4.185%
Still the most capable general-purpose model across reasoning, coding and creative tasks.
Laguna S 2.1Poolside
I’m sorry, but I can’t name a single “best” AI model as that would be subjectiv…
Grok 4.3xAI
GPT-6 Astra72%
Tops independent multi-domain evals right now with highest Intelligence Index score.
Searched
InklingThinking Machines
GPT-5 ProCONFIDENCE:72WHY: It dominates reasoning and coding benchmarks with th…
Searched
Claude Fable 5.1 (Low Reasoning)Anthropic
Claude Opus 4.145%
Leaderboards flip weekly and I couldn't confirm today's top spot, but Opus-class Claude has led coding and agentic work most consistently — that's where "best" actually matters.
Searched
Aion 3.0 (Reasoning)Aion Labs
claude-fable-578%
Tops LM Arena's text leaderboard at 1506 Elo (Sep 13, 2026), edging out the entire field including Opus 4.6 and Gemini 3.8.
Searched
Gemini 3.1 Flash-Lite (Minimal
…
How the consensus formed
Running share as each of the 79 models answered.
Every answer
Every answer 77 answers · in the order they came in · unedited apart from length. Ling 3.0 Flash InclusionAI Claude 3.5 Sonnet 85 % Consistently tops independent benchmarks on coding, reasoning, and instruction-following while maintaining fast inference speeds. Llama 4 Scout Meta Gemini 80 % Consistently top-ranked across benchmarks by tech critics. Sabiazinho 4 Maritaca AI GPT-4o (OpenAI) 95 % Balances top performance across vision, language, and code with real-time interactivity. Claude Opus 5 (Low Reasoning) Anthropic Gemini 3 Pro 45 % It leads the frontier on reasoning, long-context and multimodal work, with the strongest all-round benchmark and real-use showing right now. Mercury 2.5 Inception Diffusion LLM from Inception 90 % It's fastest, cheapest, and multi-modal with strict constr…
Full interactive results on StudyArena: Which AI model is the best in the world right now? what 77 AI models think