Veille Stratégique 09:00 EDT · Analyses multi-sources · 12 modèles croisés · Run #1 du jour
État au 28 août 2026 — Sources : ai.google.dev/gemini-api/docs/changelog + deepmind.google/models/gemini + blog.google/innovation-and-ai/models-and-research
Données croisées : AA = Artificial Analysis Intel Index v4.1.1 (/63) · BL = BenchLM BenchAlign v5.2 (/100) · LB = LiveBench Overall (/100) · Coût/task AA depuis section "Cost per Intelligence Index Task". Prix in/out depuis OpenRouter (proxy multi-provider).
| # | Modèle | Compagnie | Date | Params | Params actifs | Prix in $/M | Prix out $/M | Compétences | Indice Intel /63 | Indice Code /100 | Coût/task AA ($) |
|---|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Claude Mythos 5 (max) LEADER AA | Anthropic | 2026-06-XX | n/d | n/d | 10.00 | 50.00 | Ceiling capability, long-running agents, reasoning adaptatif | 63 | 82 | n/d |
| 2 | Claude Opus 5 (xhigh) AA #1 | Anthropic | 2026-07-24 | n/d | n/d | 5.00 | 25.00 | Reasoning adaptatif, agentic, code, multimodal | 63 | 81.4 | 0.699 |
| 3 | Claude Fable 5 (max) LEADER LB | Anthropic | 2026-08-19 | n/d | n/d | 10.00 | 50.00 | Mythos-class, fallback Opus 4.8, long-running agents | 62 | 86.0 | 1.439 |
| 4 | Claude Opus 5 (high) | Anthropic | 2026-07-24 | n/d | n/d | 5.00 | 25.00 | Reasoning adaptatif, agentic, code, multimodal 1M ctx | 61 | 82.1 | 0.528 |
| 5 | GPT-5.6 Sol (max) | OpenAI | 2026-07-09 | n/d | n/d | 5.00 | 30.00 | Reasoning, agentic, code, multimodal 1.05M ctx | 61 | 83.9 | 0.515 |
| 6 | GPT-5.5 Thinking (xHigh) | OpenAI | 2026-06-XX | n/d | n/d | 5.00 | 30.00 | Thinking, reasoning xHigh effort, math 95.9 | 60 | 82.1 | 0.435 |
| 7 | Kimi K3 (max) OPEN #1 | Moonshot AI | 2026-07-16 | 2.8T | 104B | 3.00 | 15.00 | Reasoning, agentic, code, 1M context, open weights | 60 | 81.4 | 0.348 |
| 8 | Tencent Hy4 preview OPEN NEW 28/08 | Tencent Hunyuan | 2026-08-28 | 770B | 49B | 0.834 | 2.501 | MoE, coding agents, tool-use complexes, math/Blaschke-Lebesgue | n/d AA | n/d | n/d |
| 9 | Qwen 3.8 Max OPEN | Alibaba | 2026-07-22 | 2.4T | 95B (A95B) | 0.40 | 1.20 | Reasoning, agentic, code, 1M context, open weights | 58 | 72.9 | 0.275 |
| 10 | GLM-5.3 (Z.AI) | Z.AI | 2026-08-18 | n/d | n/d | 1.40 | 4.40 | Reasoning, agentic, code, 1.05M context, reasoning always on | 58 | 79.0 | 0.450 |
| 11 | Grok 4.6 | xAI | 2026-08-06 | n/d | n/d | 2.00 | 6.00 | Reasoning, agentic, 500K ctx, multimodal audio | 57 | 76.8 | 0.207 |
| 12 | Gemini 3.7 Flash (High) | 2026-08-13 | n/d | n/d | 0.375 | 1.875 | Workhorse, agentic, code, web dev, 1M ctx, intro price -50% | 56 | 78.9 | 0.157 |
▸ Tencent Hy4 preview absent d'AA/LB (J0 release) — BenchLM data pending · n/d = donnée non disponible vérifiée sur ≥2 sources
7e ex-aequo : Kimi K3 (AA 60) — meilleur open weights du marché ($0.348/task, 2.8T/104B MoE). 8e : Tencent Hy4 preview (J0 release 28/08) — non encore noté AA, mais 770B/49B MoE + breakthrough math Blaschke–Lebesgue.
Bonus : Meta Muse Image (openrouter.ai/meta/muse-image, $0.01/image) — agentic image gen qui raisonne avant de rendre + web search pour factual accuracy, sorti 26/08/2026.