Tuesday, 25 August 2026 | Mise à jour quotidienne L'intelligence artificielle au service des constructeurs

LLM Leaderboard 2026 — AI Model Intelligence Index

In 2026, Claude Opus 4.8 tops our intelligence index at 55.7, just ahead of OpenAI’s GPT-5.5 (54.8). Google’s Gemini 3.5 Flash (50.2) and open-weight GLM-5.2 (51.1) sit close behind, with Gemini 3.1 Pro (46.5) and DeepSeek V4 Pro (44.3) rounding out the frontier. The single “best” model depends on how you weigh intelligence against price and speed.

The Intelligence column is a composite 0-100 score based on the Artificial Analysis Intelligence Index v4.1, which blends nine demanding evaluations spanning reasoning, coding, agentic tool use, and scientific knowledge. Higher is smarter — but a two-point gap at the very top rarely changes real-world output quality, so treat clusters of models as roughly equivalent rather than reading tiny differences as decisive.

Read intelligence alongside price and speed. A frontier model like Opus 4.8 costs far more per token than DeepSeek V4 Flash (40.3) or GLM-5.2, which deliver 70-90% of the intelligence at a fraction of the cost. For high-volume or latency-sensitive work, a cheaper, faster model usually wins; for the hardest reasoning and agentic tasks, the top of the table earns its premium. Sort by the metric that matches your use case.

#Modèle DéveloppeurIntelligence Contexte Entrée : $/1 million Sortie : $/1 million Open weights
1Claude Opus 5Anthropic611 million$5.00$25.00Non
2Claude Fable 5Anthropic601 million$10.00$50.00Non
3GPT-5.6 SolOpenAI591,05 million$5.00$30.00Non
4Kimi K3Moonshot AI571 million$3.00$15.00Oui
5Claude Opus 4.8Anthropic55.71 million$5.00$25.00Non
6GPT-5.5OpenAI54.81,05 million$5.00$30.00Non
7GLM 5.2Zhipu AI51.11 million$1.40$4.40Oui
8Gemini 3.5 FlashGoogle50.21 million$1.50$9.00Non
9Claude Sonnet 4.6Anthropic471 million$3.00$15.00Non
10Gemini 3.1 ProGoogle46.51,05 million$2.00$12.00Non
11DeepSeek V4-ProDeepSeek44.31 million$0.44$0.87Oui
12Kimi K2.7 CodeMoonshot AI42256 K$0.60$2.50Oui
13DeepSeek V4-FlashDeepSeek40.31 million$0.14$0.28Oui
14Claude Haiku 4.5Anthropic37200 000$1.00$5.00Non
15DeepSeek R1DeepSeek20.1128 K$0.50$2.15Oui
16Mistral Large 3Mistral AI15.9256 K$2.00$6.00Oui
17Llama 4 MaverickMeta14.31 million$0.15$0.60Oui
18Qwen3 235B-A22BAlibaba13128 K$0.45$1.80Oui
19Qwen3 32BAlibaba12128 K$0.08$0.28Oui
20Llama 4 ScoutMeta10.010 M$0.10$0.30Oui
21Gemma 3 27BGoogle7.4128 K$0.08$0.16Oui
22Phi-4Microsoft4.916 K$0.07$0.14Oui
23Claude Sonnet 5Anthropic1 million$2.00$10.00Non
24DeepSeek R1 Distill Llama 70BDeepSeek128 K$0.80$0.80Oui
25Gemini 2.5 ProGoogle (Google DeepMind)1 million (1 048 576 jetons)$1.25$10.00Non
26Gemini 3.6 FlashGoogle1 million$1.50$7.50Non
27Gemma 3 12BGoogle128 K$0.05$0.15Oui
28Gemma 3 4BGoogle128 K$0.05$0.10Oui
29Grok 4xAI256 000 jetons$3.00$15.00Non
30Kling 2.5 Turbo ProKuaishouNon
31Llama 3.1 8BMeta128 K$0.02$0.03Oui
32Llama 3.3 70BMeta128 K$0.10$0.32Oui
33Mistral 7BMistral AI32 K$0.02$0.03Oui
34Mistral NeMo 12BMistral AI128 K$0.02$0.04Oui
35NVIDIA Nemotron 3 Nano OmniNVIDIA256 KOui
36Qwen3 14BAlibaba128 K$0.12$0.24Oui
37Qwen3 30B-A3BAlibaba128 K$0.12$0.50Oui
38Qwen3 8BAlibaba128 K$0.04$0.14Oui
39Sora 2OpenAINon
40Sora 2 ProOpenAINon
41Veo 3.1GoogleNon
42Wan 2.5AlibabaNon

Click any column heading to re-sort. Intelligence = a 0–100 composite of public reasoning/knowledge benchmarks; — means not yet scored. Prices are USD per 1M tokens. Updated August 2026.

FAQ

What is the best LLM right now?

As of mid-2026, Claude Opus 4.8 is the highest-scoring model on the Artificial Analysis Intelligence Index (55.7), narrowly ahead of GPT-5.5 (54.8). Both lead on reasoning, coding, and agentic tasks, so the practical choice between them usually comes down to price, speed, and ecosystem rather than a meaningful intelligence gap.

What is the smartest open-source model?

GLM-5.2 is currently the most intelligent open-weight model, scoring 51.1 on the intelligence index — ahead of DeepSeek V4 Pro (44.3) and DeepSeek V4 Flash (40.3). That puts the best open models within roughly four points of proprietary frontier models like GPT-5.5, while remaining free to self-host or run through low-cost APIs.

Which AI model is the best value?

DeepSeek V4 Flash and GLM-5.2 offer the strongest intelligence-per-dollar. DeepSeek V4 Flash scores 40.3 at a tiny fraction of Opus 4.8’s price, delivering roughly 70% of the top score for a small percentage of the cost. For budget-sensitive, high-volume workloads, these open models are the clear value leaders.

How is the intelligence score calculated?

Our scores mirror the Artificial Analysis Intelligence Index v4.1, a 0-100 composite of nine demanding evaluations — including Humanity’s Last Exam, GPQA Diamond, Terminal-Bench, SciCode, and agentic tool-use tasks. It measures reasoning, knowledge, coding, and agentic ability in one number, so a higher score means broadly more capable across hard, real-world tasks.

Défiler vers le haut
Featured on There's An AI For That