Classement · édition du
Classement des modèles de langage, Maths
Quel est le meilleur modèle en mathématiques ?
Comment lire ce tableauComment lire ?
Le score Maths résume les résultats d'un modèle sur les benchmarks de mathématiques où il a été mesuré. L'échelle est celle du score général : 100 correspond au meilleur modèle au lancement du site, 50 à GPT-4o de fin 2024. La barre indique la marge d'erreur : deux modèles dont les barres se chevauchent sont ex æquo.
Dans l'édition du 28 septembre 2026, GPT-6 Astra (OpenAI) est premier en mathématiques avec un score Maths de 102,0, ex æquo avec 7 autres modèles dont Claude Opus 5.5 et Claude Fable 5.1.
| Rang | Empreinte | Modèle | Étiquettes | Score et marge (50 à 110) | Couverture | Semaine | Prix, € / M jetons |
|---|---|---|---|---|---|---|---|
| EX ÆQUO · RANG 1L'écart entre ces modèles est plus petit que la marge d'erreur. | |||||||
| 1 | GPT-6 AstraNOUVEAUOpenAI | 102,096,0 à 108,0 | 75 % · 6 mesures | Entrée | 8,78 € / 44 € | ||
| 2 | Claude Opus 5.5NOUVEAUAnthropic | 101,496,0 à 106,9 | 62 % · 5 mesures | Entrée | 4 € / 20 € | ||
| 3 | Claude Fable 5.1NOUVEAUAnthropic | 99,195,2 à 103,0 | 75 % · 6 mesures | Entrée | 8,78 € / 44 € | ||
| 4 | Claude Fable 5Anthropic | 98,894,9 à 102,7 | 75 % · 6 mesures | 8,78 € / 44 € | |||
| 5 | Claude Sonnet 5.5NOUVEAUAnthropic | 97,893,6 à 101,9 | 50 % · 4 mesures | Entrée | 2 € / 10 € | ||
| Sans rang | Gemini 4 ArgonPROVISOIREGoogle DeepMind | 96,588,3 à 104,6 | 12 % · 1 mesure | Entrée | Inconnu | ||
| Sans rang | AI Co-MathematicianPROVISOIREGoogle DeepMind | 96,293,7 à 98,7 | 25 % · 2 mesures | Inconnu | |||
| 6 | Claude Opus 5Anthropic | 96,092,2 à 99,7 | 62 % · 5 mesures | 4,39 € / 22 € | |||
| 7 | GPT-5.6 SolOpenAI | 96,093,1 à 98,8 | 75 % · 6 mesures | 3,51 € / 18 € | |||
| 8 | GPT-5.5 ProOpenAI | 94,992,7 à 97,1 | 62 % · 5 mesures | 26 € / 158 € | |||
| GROUPE 2Ex æquo · 8 modèles | |||||||
| 9 | GPT-5.6 TerraOpenAI | 93,391,2 à 95,5 | 62 % · 5 mesures | 2,17 € / 13 € | |||
| 10 | GPT-5.4 ProOpenAI | 92,390,7 à 93,9 | 50 % · 4 mesures | 26 € / 158 € | |||
| Sans rang | Tencent Hy4 previewPROVISOIRETencent | 92,390,0 à 94,6 | 12 % · 1 mesure | 0,74 € / 2,23 € | |||
| 11 | GPT-5.5OpenAI | 91,588,6 à 94,5 | 100 % · 8 mesures | 4,34 € / 26 € | |||
| 12 | Claude Opus 4.8Anthropic | 91,089,8 à 92,2 | 88 % · 7 mesures | 4,39 € / 22 € | |||
| 13 | GPT-5.6 LunaOpenAI | 91,088,9 à 93,1 | 62 % · 5 mesures | 0,87 € / 5,20 € | |||
| 14 | Kimi K3Moonshot | 91,087,4 à 94,6 | 62 % · 5 mesures | 2,60 € / 13 € | |||
| Sans rang | GPT-5.2 ProPROVISOIREOpenAI | 89,988,5 à 91,4 | 38 % · 3 mesures | 18 € / 147 € | |||
| 15 | Muse Spark 1.3NOUVEAUMeta AI | 89,787,8 à 91,6 | 50 % · 4 mesures | Entrée | 1,25 € / 4,25 € | ||
| 16 | Claude Sonnet 5Anthropic | 89,485,9 à 92,9 | 62 % · 5 mesures | 1,73 € / 8,67 € | |||
| GROUPE 3Ex æquo · 11 modèles | |||||||
| 17 | GPT-5.4OpenAI | 89,488,2 à 90,7 | 88 % · 7 mesures | 2,17 € / 13 € | |||
| 18 | Qwen 3.8 MaxAlibaba | 89,488,1 à 90,7 | 50 % · 4 mesures | 1,76 € / 5,27 € | |||
| Sans rang | MiMo-V2.6-ProPROVISOIREXiaomi | 88,980,7 à 97,1 | 12 % · 1 mesure | Entrée | 0,44 € / 0,87 € | ||
| 19 | Gemini 3.7 FlashGoogle DeepMind | 88,687,2 à 90,1 | 62 % · 5 mesures | 1,50 € / 7,50 € | |||
| Sans rang | DeepSeek V4.1 FlashPROVISOIREDeepSeek | 88,686,8 à 90,5 | 25 % · 2 mesures | Entrée | 0,15 € / 0,60 € | ||
| 20 | Claude Opus 4.7Anthropic | 87,986,6 à 89,1 | 88 % · 7 mesures | 4,39 € / 22 € | |||
| Sans rang | GPT-6 SolPROVISOIREOpenAI | 87,979,6 à 96,1 | 12 % · 1 mesure | Entrée | 2 € / 10 € | ||
| 21 | Grok 4.6xAI | 87,686,0 à 89,2 | 62 % · 5 mesures | 1,76 € / 5,27 € | |||
| 22 | GLM-5.3Z.ai (Zhipu AI) | 87,385,9 à 88,8 | 62 % · 5 mesures | 1,40 € / 4,40 € | |||
| 23 | Qwen3.8 Max (0902)NOUVEAUAlibaba | 87,385,1 à 89,6 | 50 % · 4 mesures | Entrée | 1,76 € / 5,27 € | ||
| 24 | Claude Opus 4.6Anthropic | 87,185,7 à 88,4 | 88 % · 7 mesures | 4,39 € / 22 € | |||
| 25 | DeepSeek V4 Flash 0731DeepSeek | 87,184,4 à 89,8 | 50 % · 4 mesures | 0,18 € / 0,35 € | |||
| 26 | DeepSeek V4 Pro 0813DeepSeek | 87,185,5 à 88,7 | 50 % · 4 mesures | 0,61 € / 1,82 € | |||
| 27 | Gemini 3.8 FlashNOUVEAUGoogle DeepMind | 86,884,8 à 88,9 | 62 % · 5 mesures | Entrée | 1,50 € / 7,50 € | ||
| Sans rang | Muse Spark 1.2PROVISOIREMeta AI | 86,885,0 à 88,6 | 12 % · 1 mesure | 1,25 € / 4,25 € | |||
| Sans rang | Muse Spark 1.1PROVISOIREMeta AI | 86,384,3 à 88,2 | 25 % · 2 mesures | 1,10 € / 3,73 € | |||
| GROUPE 4Ex æquo · 18 modèles | |||||||
| 28 | GLM-5.2Z.ai (Zhipu AI) | 85,583,1 à 87,9 | 62 % · 5 mesures | 0,98 € / 3,08 € | |||
| 29 | Qwen3.7-MaxAlibaba | 85,583,3 à 87,7 | 62 % · 5 mesures | 1,08 € / 3,25 € | |||
| 30 | Gemini 3.6 FlashGoogle DeepMind | 85,283,7 à 86,8 | 62 % · 5 mesures | 1,30 € / 6,50 € | |||
| Sans rang | Grok 4.7PROVISOIRExAI | 85,281,9 à 88,6 | 25 % · 2 mesures | Entrée | 1,76 € / 5,27 € | ||
| 31 | Gemini 3.5 FlashGoogle DeepMind | 85,083,7 à 86,3 | 88 % · 7 mesures | 1,30 € / 7,80 € | |||
| 32 | Grok 4.5xAI | 85,083,4 à 86,5 | 62 % · 5 mesures | 1,76 € / 5,27 € | |||
| 33 | GPT-5.2OpenAI | 84,782,4 à 87,1 | 88 % · 7 mesures | 1,52 € / 12 € | |||
| 34 | Gemini 3.1 ProGoogle DeepMind | 84,783,3 à 86,1 | 88 % · 7 mesures | 2 € / 12 € | |||
| Sans rang | ERNIE 5.1PROVISOIREBaidu | 84,776,5 à 93,0 | 12 % · 1 mesure | 0,66 € / 2,63 € | |||
| Sans rang | Hy3PROVISOIRETencent | 84,776,5 à 93,0 | 12 % · 1 mesure | 0,15 € / 0,50 € | |||
| 35 | Claude Sonnet 4.6Anthropic | 84,580,4 à 88,5 | 62 % · 5 mesures | 2,63 € / 13 € | |||
| 36 | Gemini 3 ProGoogle DeepMind | 83,981,8 à 86,1 | 62 % · 5 mesures | 1,73 € / 10 € | |||
| Sans rang | GPT-5 ProPROVISOIREOpenAI | 83,982,0 à 85,9 | 38 % · 3 mesures | 13 € / 105 € | |||
| Sans rang | Qwen3.5 Max PreviewPROVISOIREAlibaba | 83,975,7 à 92,2 | 12 % · 1 mesure | Inconnu | |||
| 37 | Kimi K2.6Moonshot | 83,781,8 à 85,5 | 88 % · 7 mesures | 0,59 € / 2,96 € | |||
| Sans rang | GPT-6 LunaPROVISOIREOpenAI | 83,475,2 à 91,6 | 12 % · 1 mesure | Entrée | 0,10 € / 0,50 € | ||
| 38 | GLM-5.3-FlashZ.ai (Zhipu AI) | 83,281,3 à 85,0 | 62 % · 5 mesures | 0,15 € / 0,50 € | |||
| 39 | Muse SparkMeta AI | 83,280,7 à 85,6 | 62 % · 5 mesures | Inconnu | |||
| Sans rang | MiMo-V2.5-ProPROVISOIREXiaomi | 83,280,6 à 85,8 | 25 % · 2 mesures | 0,38 € / 0,76 € | |||
| 40 | GPT-5OpenAI | 82,681,0 à 84,3 | 88 % · 7 mesures | 1,08 € / 8,67 € | |||
| Sans rang | GPT-5.2 ChatPROVISOIREOpenAI | 82,674,4 à 90,9 | 12 % · 1 mesure | 1,54 € / 12 € | |||
| Sans rang | Grok 4.20 Beta 1PROVISOIRExAI | 82,674,4 à 90,9 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Grok 4.20 Multi Agent BetaPROVISOIRExAI | 82,674,4 à 90,9 | 12 % · 1 mesure | 1,10 € / 2,19 € | |||
| Sans rang | Kimi K2.7 CodePROVISOIREMoonshot | 82,479,8 à 85,0 | 38 % · 3 mesures | 0,83 € / 3,51 € | |||
| Sans rang | MiMo-V2.6-FlashPROVISOIREXiaomi | 82,173,9 à 90,3 | 12 % · 1 mesure | Entrée | 0,14 € / 0,28 € | ||
| 41 | Gemini 3 FlashGoogle DeepMind | 81,979,9 à 83,9 | 88 % · 7 mesures | 0,43 € / 2,60 € | |||
| 42 | GLM-5.1Z.ai (Zhipu AI) | 81,679,0 à 84,2 | 75 % · 6 mesures | 0,85 € / 2,67 € | |||
| 43 | GPT-5.1OpenAI | 81,678,9 à 84,3 | 50 % · 4 mesures | 1,08 € / 8,67 € | |||
| Sans rang | Qwen 3.8 27BPROVISOIREAlibaba | 81,678,6 à 84,6 | 25 % · 2 mesures | 0,35 € / 2,75 € | |||
| Sans rang | MiMo-V2.5PROVISOIREXiaomi | 81,377,3 à 85,4 | 25 % · 2 mesures | 0,12 € / 0,25 € | |||
| 44 | Grok 4.20xAI | 81,178,6 à 83,5 | 50 % · 4 mesures | 1,08 € / 2,17 € | |||
| Sans rang | Gemini 2.5 Deep ThinkPROVISOIREGoogle DeepMind | 80,878,4 à 83,3 | 25 % · 2 mesures | Inconnu | |||
| 45 | DeepSeek-V4-ProDeepSeek | 80,677,8 à 83,3 | 62 % · 5 mesures | 0,38 € / 0,75 € | |||
| GROUPE 5Ex æquo · 13 modèles | |||||||
| 46 | GPT-5.4 MiniOpenAI | 80,678,4 à 82,7 | 88 % · 7 mesures | 0,65 € / 3,90 € | |||
| Sans rang | Grok 4.20 Beta (0309)PROVISOIRExAI | 80,672,3 à 88,8 | 12 % · 1 mesure | 1,76 € / 5,27 € | |||
| 47 | Inkling-SmallThinking Machines | 80,377,2 à 83,4 | 62 % · 5 mesures | 0,45 € / 1,20 € | |||
| Sans rang | DeepSeek-V4-FlashPROVISOIREDeepSeek | 80,372,1 à 88,5 | 12 % · 1 mesure | 0,09 € / 0,17 € | |||
| Sans rang | Seed 2.0 ProPROVISOIREByteDance | 80,372,1 à 88,5 | 12 % · 1 mesure | 0,44 € / 2,63 € | |||
| 48 | Grok 4.3 BetaxAI | 80,077,3 à 82,7 | 62 % · 5 mesures | 1,08 € / 2,17 € | |||
| Sans rang | MiMo-V2-ProPROVISOIREXiaomi | 80,071,8 à 88,3 | 12 % · 1 mesure | 0,38 € / 0,76 € | |||
| 49 | GPT-5 miniOpenAI | 79,577,3 à 81,7 | 88 % · 7 mesures | 0,22 € / 1,73 € | |||
| 50 | Kimi K2.5Moonshot | 79,576,7 à 82,3 | 50 % · 4 mesures | 0,35 € / 1,65 € | |||
| 51 | GPT-5.4 NanoOpenAI | 79,076,9 à 81,0 | 88 % · 7 mesures | 0,17 € / 1,08 € | |||
| 52 | Qwen 3.6 PlusAlibaba | 79,076,8 à 81,1 | 62 % · 5 mesures | 0,28 € / 1,69 € | |||
| Sans rang | GPT-5.1-Codex-MaxPROVISOIREOpenAI | 79,074,2 à 83,8 | 12 % · 1 mesure | 1,10 € / 8,78 € | |||
| 53 | Claude Opus 4.5Anthropic | 78,574,1 à 82,8 | 88 % · 7 mesures | 4,39 € / 22 € | |||
| Sans rang | GPT-5.3 ChatPROVISOIREOpenAI | 78,270,0 à 86,4 | 12 % · 1 mesure | 1,52 € / 12 € | |||
| 54 | Qwen 3.6 Max (Preview)Alibaba | 77,975,1 à 80,7 | 50 % · 4 mesures | 1,14 € / 6,85 € | |||
| Sans rang | ERNIE 5.0PROVISOIREBaidu | 77,969,7 à 86,2 | 12 % · 1 mesure | 0,74 € / 2,96 € | |||
| 55 | o4-miniOpenAI | 77,774,9 à 80,5 | 75 % · 6 mesures | 0,95 € / 3,82 € | |||
| Sans rang | Kimi K2.5 InstantPROVISOIREMoonshot | 77,769,5 à 85,9 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen3.7-PlusPROVISOIREAlibaba | 77,773,4 à 82,0 | 38 % · 3 mesures | 0,28 € / 1,11 € | |||
| Sans rang | Mistral Medium 3.5PROVISOIREMistral AI | 77,470,9 à 83,9 | 25 % · 2 mesures | 1,30 € / 6,50 € | |||
| 56 | DeepSeek-V3.2DeepSeek | 77,274,5 à 79,8 | 62 % · 5 mesures | 0,20 € / 0,30 € | |||
| Sans rang | Qwen3.6 27BPROVISOIREAlibaba | 77,273,8 à 80,5 | 25 % · 2 mesures | 0,53 € / 3,16 € | |||
| Sans rang | GLM-5V-TurboPROVISOIREZ.ai (Zhipu AI) | 76,968,7 à 85,1 | 12 % · 1 mesure | 1,05 € / 3,51 € | |||
| 57 | InklingThinking Machines | 76,673,4 à 79,8 | 62 % · 5 mesures | 0,95 € / 4,05 € | |||
| Sans rang | Muse Glimmer 30BPROVISOIREMeta AI | 76,668,4 à 84,9 | 12 % · 1 mesure | 0,30 € / 1,10 € | |||
| Sans rang | MiMo V2 OmniPROVISOIREXiaomi | 76,468,1 à 84,6 | 12 % · 1 mesure | 0,12 € / 0,25 € | |||
| Sans rang | Qwen 3.5 Plus (hosted 397B-A17B)PROVISOIREAlibaba | 76,473,3 à 79,4 | 38 % · 3 mesures | 0,35 € / 2,11 € | |||
| 58 | Kimi K2 ThinkingMoonshot | 76,172,3 à 80,0 | 50 % · 4 mesures | 0,52 € / 2,17 € | |||
| Sans rang | Qwen3.5 397B-A17BPROVISOIREAlibaba | 76,172,6 à 79,7 | 38 % · 3 mesures | 0,34 € / 2,03 € | |||
| GROUPE 6Ex æquo · 13 modèles | |||||||
| 59 | o3OpenAI | 75,973,5 à 78,2 | 62 % · 5 mesures | 1,76 € / 7,02 € | |||
| 60 | Grok 4xAI | 75,672,5 à 78,7 | 50 % · 4 mesures | 2,63 € / 13 € | |||
| Sans rang | Nemotron 3 UltraPROVISOIRENVIDIA | 75,370,2 à 80,4 | 38 % · 3 mesures | Inconnu | |||
| Sans rang | Grok 4.1PROVISOIRExAI | 74,866,6 à 83,0 | 12 % · 1 mesure | 1,76 € / 8,78 € | |||
| Sans rang | Hunyuan Hy3 PreviewPROVISOIRETencent | 74,566,3 à 82,8 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | MiniMax-M2.7PROVISOIREMiniMax | 74,567,0 à 82,1 | 25 % · 2 mesures | 0,24 € / 1,04 € | |||
| Sans rang | Grok 4.1 FastPROVISOIRExAI | 74,366,6 à 82,0 | 25 % · 2 mesures | 0,17 € / 0,43 € | |||
| Sans rang | Qwen3.5-122B-A10BPROVISOIREAlibaba | 74,366,1 à 82,5 | 12 % · 1 mesure | 0,35 € / 2,81 € | |||
| 61 | Claude Sonnet 4.5Anthropic | 74,070,3 à 77,7 | 88 % · 7 mesures | 2,60 € / 13 € | |||
| 62 | GLM-5Z.ai (Zhipu AI) | 74,070,9 à 77,2 | 50 % · 4 mesures | 0,52 € / 1,66 € | |||
| Sans rang | ChatGPT-4o LatestPROVISOIREOpenAI | 74,065,8 à 82,2 | 12 % · 1 mesure | 4,39 € / 18 € | |||
| Sans rang | gpt-oss-120bPROVISOIREOpenAI | 74,065,9 à 82,2 | 25 % · 2 mesures | 0,03 € / 0,16 € | |||
| 63 | Gemini 2.5 Pro (Jun 2025)Google DeepMind | 73,871,3 à 76,2 | 75 % · 6 mesures | 1,10 € / 8,78 € | |||
| Sans rang | Gemma 4 26B A4BPROVISOIREGoogle DeepMind | 73,867,0 à 80,5 | 25 % · 2 mesures | 0,05 € / 0,29 € | |||
| Sans rang | Grok 4 Fast ChatPROVISOIRExAI | 73,865,5 à 82,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mercury 2.5PROVISOIREInception Labs | 73,865,6 à 82,0 | 12 % · 1 mesure | Entrée | 0,04 € / 0,13 € | ||
| Sans rang | Qwen3.5-27BPROVISOIREAlibaba | 73,865,5 à 82,0 | 12 % · 1 mesure | 0,26 € / 2,11 € | |||
| Sans rang | Gemini 3.1 Flash-LitePROVISOIREGoogle DeepMind | 73,569,8 à 77,2 | 38 % · 3 mesures | 0,22 € / 1,30 € | |||
| Sans rang | DeepSeek-V3.2-ExpPROVISOIREDeepSeek | 73,064,8 à 81,2 | 12 % · 1 mesure | 0,25 € / 0,38 € | |||
| Sans rang | Qwen 3.6 35B-A3BPROVISOIREAlibaba | 73,069,0 à 76,9 | 25 % · 2 mesures | 0,22 € / 1,30 € | |||
| Sans rang | Qwen3.7 FlashPROVISOIREAlibaba | 72,768,5 à 77,0 | 25 % · 2 mesures | 0,03 € / 0,11 € | |||
| Sans rang | Solar Pro 4PROVISOIREUpstage | 72,564,2 à 80,7 | 12 % · 1 mesure | 0,26 € / 1,05 € | |||
| 64 | Qwen 3.6 FlashAlibaba | 72,269,4 à 75,0 | 50 % · 4 mesures | 0,16 € / 0,99 € | |||
| Sans rang | MiniMax-M2.5PROVISOIREMiniMax | 72,263,5 à 80,9 | 25 % · 2 mesures | 0,13 € / 1 € | |||
| Sans rang | DeepSeek-V3.1PROVISOIREDeepSeek | 71,963,7 à 80,2 | 12 % · 1 mesure | 0,18 € / 0,69 € | |||
| 65 | Gemini 3.5 Flash-LiteGoogle DeepMind | 71,766,4 à 77,0 | 62 % · 5 mesures | 0,26 € / 2,17 € | |||
| 66 | Qwen3-235B-A22B-Thinking (Jul 2025)Alibaba | 71,767,6 à 75,7 | 50 % · 4 mesures | 0,26 € / 2,55 € | |||
| Sans rang | Grok 4 HeavyPROVISOIRExAI | 71,763,7 à 79,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Kimi K2 (Sep 2025)PROVISOIREMoonshot | 71,763,5 à 79,9 | 12 % · 1 mesure | 0,55 € / 2,22 € | |||
| 67 | o3-miniOpenAI | 71,468,6 à 74,2 | 75 % · 6 mesures | 0,97 € / 3,86 € | |||
| Sans rang | Qwen3 VL 235B A22B InstructPROVISOIREAlibaba | 71,463,2 à 79,6 | 12 % · 1 mesure | 0,47 € / 2,34 € | |||
| Sans rang | DeepSeek-V3.1-TerminusPROVISOIREDeepSeek | 70,962,7 à 79,1 | 12 % · 1 mesure | 0,24 € / 0,88 € | |||
| Sans rang | GLM-4.5PROVISOIREZ.ai (Zhipu AI) | 70,962,7 à 79,1 | 12 % · 1 mesure | 0,52 € / 1,91 € | |||
| Sans rang | Grok 4 FastPROVISOIRExAI | 70,962,7 à 79,1 | 12 % · 1 mesure | 0,17 € / 0,43 € | |||
| 68 | GPT-5 nanoOpenAI | 70,666,6 à 74,7 | 88 % · 7 mesures | 0,04 € / 0,35 € | |||
| 69 | GPT-5.5 InstantOpenAI | 70,665,1 à 76,2 | 50 % · 4 mesures | 4,39 € / 26 € | |||
| 70 | Qwen 3.5 Flash (hosted 35B-A3B)Alibaba | 70,667,1 à 74,1 | 62 % · 5 mesures | 0,09 € / 0,35 € | |||
| Sans rang | Laguna M.1PROVISOIREPoolside | 70,664,7 à 76,6 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | LongCat Flash ChatPROVISOIREMeituan | 70,662,4 à 78,9 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemma 4 31B ITPROVISOIREGoogle DeepMind | 70,461,7 à 79,1 | 25 % · 2 mesures | 0,10 € / 0,31 € | |||
| Sans rang | MiniMax-M2.1PROVISOIREMiniMax | 70,462,1 à 78,6 | 12 % · 1 mesure | 0,26 € / 1,05 € | |||
| Sans rang | Qwen3-235B-A22B-Instruct (Jul 2025)PROVISOIREAlibaba | 70,462,1 à 78,6 | 12 % · 1 mesure | 0,19 € / 0,77 € | |||
| 71 | GLM-4.7Z.ai (Zhipu AI) | 70,164,1 à 76,1 | 62 % · 5 mesures | 0,35 € / 1,52 € | |||
| Sans rang | Gemini 2.5 Flash (Sep 2025)PROVISOIREGoogle DeepMind | 70,161,9 à 78,3 | 12 % · 1 mesure | 0,26 € / 2,19 € | |||
| Sans rang | Laguna XS.2PROVISOIREPoolside | 70,164,7 à 75,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | MiniMax-M3PROVISOIREMiniMax | 70,160,1 à 80,1 | 38 % · 3 mesures | 0,24 € / 0,97 € | |||
| Sans rang | Qwen3-MaxPROVISOIREAlibaba | 70,166,5 à 73,7 | 38 % · 3 mesures | 0,68 € / 3,38 € | |||
| Sans rang | Grok-3 miniPROVISOIRExAI | 69,665,7 à 73,5 | 38 % · 3 mesures | 0,26 € / 0,43 € | |||
| GROUPE 7Ex æquo · 6 modèles | |||||||
| 72 | o1OpenAI | 69,366,5 à 72,2 | 50 % · 4 mesures | 13 € / 53 € | |||
| Sans rang | chutes/Qwen3-Next-80B-A3B-InstructPROVISOIREAlibaba | 69,361,1 à 77,5 | 12 % · 1 mesure | 0,44 € / 1,76 € | |||
| Sans rang | Step 3.5 FlashPROVISOIREStepFun | 69,361,1 à 77,5 | 12 % · 1 mesure | 0,08 € / 0,26 € | |||
| Sans rang | Gemini 2.5 Flash (Apr 2025)PROVISOIREGoogle DeepMind | 69,165,4 à 72,7 | 12 % · 1 mesure | 0,13 € / 0,53 € | |||
| Sans rang | Hunyuan-T1PROVISOIRETencent | 68,860,6 à 77,0 | 12 % · 1 mesure | Inconnu | |||
| 73 | Claude Opus 4.1Anthropic | 68,565,4 à 71,7 | 75 % · 6 mesures | 13 € / 66 € | |||
| Sans rang | Gemini 2.5 Flash (May 2025)PROVISOIREGoogle DeepMind | 68,565,0 à 72,1 | 12 % · 1 mesure | 0,12 € / 2,77 € | |||
| Sans rang | Qwen3 VL 235B A22B ThinkingPROVISOIREAlibaba | 68,560,3 à 76,8 | 12 % · 1 mesure | 0,61 € / 2,46 € | |||
| Sans rang | Qwen3.5-35B-A3BPROVISOIREAlibaba | 68,364,8 à 71,7 | 25 % · 2 mesures | 0,12 € / 0,87 € | |||
| 74 | Claude Sonnet 4Anthropic | 68,064,6 à 71,4 | 50 % · 4 mesures | 2,60 € / 13 € | |||
| Sans rang | Qwen3-30B-A3B-Thinking (Jul 2025)PROVISOIREAlibaba | 68,064,5 à 71,5 | 12 % · 1 mesure | 0,32 € / 1,14 € | |||
| Sans rang | Kimi K2 (Jul 2025)PROVISOIREMoonshot | 67,859,5 à 76,0 | 12 % · 1 mesure | 0,50 € / 2,02 € | |||
| Sans rang | Mistral Large 3PROVISOIREMistral AI | 67,859,5 à 76,0 | 12 % · 1 mesure | 0,43 € / 1,30 € | |||
| Sans rang | Gemini 2.5 Flash (Jun 2025)PROVISOIREGoogle DeepMind | 67,560,4 à 74,5 | 38 % · 3 mesures | 0,26 € / 2,17 € | |||
| 75 | Claude Haiku 4.5Anthropic | 67,264,0 à 70,5 | 50 % · 4 mesures | 0,88 € / 4,39 € | |||
| Sans rang | DeepSeek-R1 (May 2025)PROVISOIREDeepSeek | 67,263,1 à 71,4 | 25 % · 2 mesures | 0,43 € / 1,86 € | |||
| Sans rang | GLM-4.6PROVISOIREZ.ai (Zhipu AI) | 67,259,9 à 74,6 | 38 % · 3 mesures | 0,37 € / 1,51 € | |||
| Sans rang | MiMo-V2-FlashPROVISOIREXiaomi | 67,259,0 à 75,5 | 12 % · 1 mesure | 0,12 € / 0,25 € | |||
| Sans rang | Mistral Medium 3.1PROVISOIREMistral AI | 67,259,0 à 75,5 | 12 % · 1 mesure | 0,35 € / 1,73 € | |||
| Sans rang | Qwen3-235B-A22BPROVISOIREAlibaba | 67,259,0 à 75,5 | 12 % · 1 mesure | 0,61 € / 2,46 € | |||
| Sans rang | seed-oss-36b-instructPROVISOIREByteDance | 67,263,9 à 70,6 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen3-14BPROVISOIREAlibaba | 67,063,7 à 70,3 | 12 % · 1 mesure | 0,31 € / 1,23 € | |||
| Sans rang | Qwen3-32BPROVISOIREAlibaba | 67,063,7 à 70,2 | 25 % · 2 mesures | 0,07 € / 0,24 € | |||
| 76 | Claude Opus 4Anthropic | 66,762,5 à 71,0 | 50 % · 4 mesures | 13 € / 66 € | |||
| Sans rang | Qwen3-Coder-480B-A35BPROVISOIREAlibaba | 66,258,0 à 74,4 | 12 % · 1 mesure | 0,19 € / 1,56 € | |||
| Sans rang | Trinity Large ThinkingPROVISOIREArcee AI | 66,258,0 à 74,4 | 12 % · 1 mesure | 0,19 € / 0,74 € | |||
| Sans rang | gpt-oss-20bPROVISOIREOpenAI | 65,959,0 à 72,9 | 25 % · 2 mesures | 0,03 € / 0,12 € | |||
| Sans rang | Qwen3-30B-A3B-Instruct (Jul 2025)PROVISOIREAlibaba | 65,762,4 à 68,9 | 25 % · 2 mesures | 0,26 € / 0,44 € | |||
| Sans rang | Qwen3.5-9BPROVISOIREAlibaba | 65,762,5 à 68,8 | 12 % · 1 mesure | 0,10 € / 0,15 € | |||
| Sans rang | Qwen3-30B-A3BPROVISOIREAlibaba | 65,460,1 à 70,7 | 25 % · 2 mesures | 0,10 € / 0,43 € | |||
| Sans rang | Qwen3-Next-80B-A3B ThinkingPROVISOIREAlibaba | 64,956,7 à 73,1 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Claude 3.7 SonnetPROVISOIREAnthropic | 64,661,6 à 67,7 | 38 % · 3 mesures | 2,63 € / 13 € | |||
| Sans rang | GLM-4.5-AirPROVISOIREZ.ai (Zhipu AI) | 64,656,4 à 72,9 | 12 % · 1 mesure | 0,18 € / 0,97 € | |||
| Sans rang | QwQ-32BPROVISOIREAlibaba | 64,661,0 à 68,3 | 25 % · 2 mesures | 0,13 € / 0,35 € | |||
| Sans rang | Gemini 2.0 Flash Thinking (Jan 2025)PROVISOIREGoogle DeepMind | 64,461,3 à 67,4 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | glm-4.7-flashPROVISOIREZ.ai (Zhipu AI) | 64,460,8 à 67,9 | 25 % · 2 mesures | Inconnu | |||
| 77 | Grok 3xAI | 63,860,9 à 66,7 | 50 % · 4 mesures | 2,63 € / 13 € | |||
| Sans rang | DeepSeek-R1-Distill-Qwen-32BPROVISOIREDeepSeek | 63,860,8 à 66,9 | 12 % · 1 mesure | 0,25 € / 0,76 € | |||
| Sans rang | Qwen3-8BPROVISOIREAlibaba | 63,860,8 à 66,9 | 12 % · 1 mesure | 0,04 € / 0,35 € | |||
| Sans rang | Qwen3.5-4BPROVISOIREAlibaba | 63,860,8 à 66,9 | 12 % · 1 mesure | 0,04 € / 0,06 € | |||
| Sans rang | DeepSeek-R1PROVISOIREDeepSeek | 63,360,3 à 66,3 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemini 2.5 Flash-Lite (Sep 2025)PROVISOIREGoogle DeepMind | 63,154,8 à 71,3 | 12 % · 1 mesure | 0,09 € / 0,35 € | |||
| Sans rang | Qwen Plus (Apr 2025)PROVISOIREAlibaba | 63,155,1 à 71,1 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | qwen3-4b-instruct-2507PROVISOIREAlibaba | 62,859,8 à 65,8 | 12 % · 1 mesure | 0,18 € / 0,18 € | |||
| Sans rang | DeepSeek-R1-Distill-Llama-70BPROVISOIREDeepSeek | 62,559,6 à 65,5 | 12 % · 1 mesure | 0,61 € / 0,69 € | |||
| Sans rang | DeepSeek-R1-Distill-Qwen-14BPROVISOIREDeepSeek | 62,359,3 à 65,3 | 12 % · 1 mesure | 0,13 € / 0,38 € | |||
| Sans rang | Llama 3.1 Nemotron Ultra 253BPROVISOIRENVIDIA | 62,354,1 à 70,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Llama 3.3 Nemotron Super 49B v1.5PROVISOIRENVIDIA | 62,354,1 à 70,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | MiniMax-M2PROVISOIREMiniMax | 62,354,1 à 70,5 | 12 % · 1 mesure | 0,22 € / 0,87 € | |||
| Sans rang | MiniMax-M1-80kPROVISOIREMiniMax | 62,053,8 à 70,2 | 12 % · 1 mesure | 0,12 € / 1,10 € | |||
| Sans rang | Gemini 2.5 Flash-Lite (Jun 2025)PROVISOIREGoogle DeepMind | 61,853,5 à 70,0 | 12 % · 1 mesure | 0,09 € / 0,35 € | |||
| Sans rang | Nemotron 3 Super 120BPROVISOIRENVIDIA | 61,853,5 à 70,0 | 12 % · 1 mesure | 0,08 € / 0,39 € | |||
| Sans rang | INTELLECT-3PROVISOIREPrime Intellect | 61,553,3 à 69,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | o1-miniPROVISOIREOpenAI | 61,558,4 à 64,6 | 38 % · 3 mesures | 0,97 € / 3,86 € | |||
| GROUPE 8Ex æquo · 3 modèles | |||||||
| 78 | GPT-4.1 miniOpenAI | 61,258,0 à 64,5 | 50 % · 4 mesures | 0,35 € / 1,39 € | |||
| Sans rang | Hunyuan-TurboSPROVISOIRETencent | 61,253,0 à 69,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | GLM-4.5VPROVISOIREZ.ai (Zhipu AI) | 60,252,0 à 68,4 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Nemotron 3.5 LightningPROVISOIRENVIDIA | 59,951,7 à 68,2 | 12 % · 1 mesure | 0,08 € / 0,20 € | |||
| Sans rang | Step 3PROVISOIREStepFun | 59,951,7 à 68,2 | 12 % · 1 mesure | Inconnu | |||
| 79 | GPT-4.1OpenAI | 59,755,0 à 64,3 | 62 % · 5 mesures | 1,76 € / 7,02 € | |||
| Sans rang | deepseek-r1-0528-qwen3-8bPROVISOIREDeepSeek | 59,756,6 à 62,8 | 12 % · 1 mesure | 0,05 € / 0,08 € | |||
| Sans rang | GPT-4.5PROVISOIREOpenAI | 59,752,1 à 67,3 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | DeepSeek-V3 (Mar 2025)PROVISOIREDeepSeek | 59,155,1 à 63,2 | 25 % · 2 mesures | 0,17 € / 0,67 € | |||
| Sans rang | o1-previewPROVISOIREOpenAI | 57,651,8 à 63,3 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Mistral Medium 3PROVISOIREMistral AI | 57,354,0 à 60,7 | 38 % · 3 mesures | 0,35 € / 1,76 € | |||
| Sans rang | Gemini 2.0 Flash (Feb 2025)PROVISOIREGoogle DeepMind | 57,153,3 à 60,8 | 38 % · 3 mesures | 0,09 € / 0,37 € | |||
| Sans rang | Ling-flash-2.0PROVISOIREAnt Group | 56,848,6 à 65,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Magistral Small 1.0PROVISOIREMistral AI | 56,553,0 à 60,0 | 12 % · 1 mesure | 0,43 € / 1,30 € | |||
| Sans rang | Mistral Small 3.2PROVISOIREMistral AI | 56,553,1 à 60,0 | 25 % · 2 mesures | 0,07 € / 0,17 € | |||
| Sans rang | Magistral Small 1.2PROVISOIREMistral AI | 55,852,1 à 59,4 | 12 % · 1 mesure | 0,44 € / 1,32 € | |||
| Sans rang | GPT-4.1 nanoPROVISOIREOpenAI | 55,248,2 à 62,2 | 38 % · 3 mesures | 0,09 € / 0,35 € | |||
| Sans rang | Nemotron 3 Nano 30BPROVISOIRENVIDIA | 55,046,7 à 63,2 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | glm-4.7-flash_nonePROVISOIREZ.ai (Zhipu AI) | 54,750,8 à 58,6 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemini 1.5 Pro (Sept 2024)PROVISOIREGoogle DeepMind | 54,450,3 à 58,6 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Gemini 2.0 Flash-LitePROVISOIREGoogle DeepMind | 54,246,0 à 62,4 | 12 % · 1 mesure | 0,07 € / 0,26 € | |||
| Sans rang | Ring-flash-2.0PROVISOIREAnt Group | 54,246,0 à 62,4 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Cohere Command APROVISOIRECohere | 53,945,7 à 62,1 | 12 % · 1 mesure | 2,17 € / 8,67 € | |||
| Sans rang | Gemma 3 27BPROVISOIREGoogle DeepMind | 53,949,9 à 57,9 | 25 % · 2 mesures | 0,07 € / 0,14 € | |||
| Sans rang | Llama 4 MaverickPROVISOIREMeta AI | 53,148,9 à 57,4 | 38 % · 3 mesures | 0,13 € / 0,52 € | |||
| Sans rang | Granite 4.1 8BPROVISOIREIBM | 52,644,4 à 60,8 | 12 % · 1 mesure | 0,04 € / 0,09 € | |||
| Sans rang | Qwen2.5-MaxPROVISOIREAlibaba | 52,445,3 à 59,4 | 38 % · 3 mesures | 1,40 € / 5,62 € | |||
| Sans rang | Qwen Plus (Jan 2025)PROVISOIREAlibaba | 52,147,4 à 56,8 | 25 % · 2 mesures | 0,35 € / 1,05 € | |||
| Sans rang | DeepSeek-V3PROVISOIREDeepSeek | 51,345,2 à 57,4 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Yi-LightningPROVISOIRE01.AI | 51,343,1 à 59,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemma 3 12BPROVISOIREGoogle DeepMind | 51,146,1 à 56,0 | 25 % · 2 mesures | 0,04 € / 0,13 € | |||
| Sans rang | Grok-2 (Aug 2024)PROVISOIRExAI | 51,142,8 à 59,3 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemini 1.5 Flash (Sep 2024)PROVISOIREGoogle DeepMind | 50,545,0 à 56,1 | 38 % · 3 mesures | Inconnu | |||
| Sans rang | OLMo 3.1 32B InstructPROVISOIREAllen Institute for AI | 50,542,3 à 58,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | DeepSeek-V2.5 (Dec 2024)PROVISOIREDeepSeek | 50,041,8 à 58,2 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | GPT-4 Turbo (Nov 2023)PROVISOIREOpenAI | 49,541,3 à 57,7 | 12 % · 1 mesure | 8,78 € / 26 € | |||
| Sans rang | Grok-2 (Dec 2024)PROVISOIRExAI | 49,242,6 à 55,9 | 25 % · 2 mesures | Inconnu | |||
| 80 | Claude 3.5 SonnetAnthropic | 48,739,3 à 58,1 | 50 % · 4 mesures | 2,63 € / 13 € | |||
| Sans rang | Llama 3.1-405BPROVISOIREMeta AI | 48,742,1 à 55,3 | 25 % · 2 mesures | 2,38 € / 2,60 € | |||
| Sans rang | Claude 3.5 Sonnet (October 2024)PROVISOIREAnthropic | 48,238,2 à 58,1 | 38 % · 3 mesures | Inconnu | |||
| Sans rang | Magistral Medium 1.1PROVISOIREMistral AI | 48,240,0 à 56,4 | 12 % · 1 mesure | 1,73 € / 4,34 € | |||
| Sans rang | OLMo 3 32B ThinkPROVISOIREAi2 | 47,939,7 à 56,1 | 12 % · 1 mesure | 0,13 € / 0,43 € | |||
| Sans rang | Phi-4PROVISOIREMicrosoft | 47,940,6 à 55,2 | 25 % · 2 mesures | 0,06 € / 0,12 € | |||
| Sans rang | Llama 4 ScoutPROVISOIREMeta AI | 47,440,4 à 54,4 | 38 % · 3 mesures | 0,09 € / 0,26 € | |||
| Sans rang | DeepSeek-V2.5 (Sep 2024)PROVISOIREDeepSeek | 47,138,9 à 55,3 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | GPT-4o (May 2024)PROVISOIREOpenAI | 46,939,5 à 54,3 | 25 % · 2 mesures | 4,39 € / 13 € | |||
| Sans rang | Mistral Large 2 (Jul 2024)PROVISOIREMistral AI | 46,940,3 à 53,4 | 25 % · 2 mesures | 1,76 € / 5,27 € | |||
| Sans rang | Qwen2.5-72BPROVISOIREAlibaba | 46,940,2 à 53,6 | 25 % · 2 mesures | 2,46 € / 7,37 € | |||
| GROUPE 9Un seul modèle dans ce groupe | |||||||
| 81 | GPT-4o (Aug 2024)OpenAI | 46,639,1 à 54,1 | 50 % · 4 mesures | 2,19 € / 8,78 € | |||
| Sans rang | Llama-3.1-Nemotron-70B-InstructPROVISOIRENVIDIA | 46,638,4 à 54,8 | 12 % · 1 mesure | 0,52 € / 0,52 € | |||
| Sans rang | Gemini 1.5 Pro (May 2024)PROVISOIREGoogle DeepMind | 46,439,3 à 53,4 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | GPT-4 Turbo (Apr 2024)PROVISOIREOpenAI | 46,439,3 à 53,5 | 38 % · 3 mesures | Inconnu | |||
| Sans rang | Mistral Large 2 (Nov 2024)PROVISOIREMistral AI | 46,439,5 à 53,2 | 38 % · 3 mesures | 1,76 € / 5,27 € | |||
| Sans rang | Claude 3 OpusPROVISOIREAnthropic | 45,837,8 à 53,9 | 25 % · 2 mesures | 13 € / 66 € | |||
| Sans rang | Gemma 3n E4BPROVISOIREGoogle DeepMind | 45,837,6 à 54,0 | 12 % · 1 mesure | 0,05 € / 0,10 € | |||
| Sans rang | GPT-4o (Nov 2024)PROVISOIREOpenAI | 45,837,8 à 53,8 | 25 % · 2 mesures | 2,19 € / 8,78 € | |||
| Sans rang | Qwen2.5-32BPROVISOIREAlibaba | 45,838,4 à 53,2 | 12 % · 1 mesure | 0,61 € / 2,46 € | |||
| Sans rang | Grok-2 miniPROVISOIRExAI | 45,637,4 à 53,8 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Llama 3.3 70BPROVISOIREMeta AI | 45,638,0 à 53,1 | 25 % · 2 mesures | 0,09 € / 0,28 € | |||
| Sans rang | Mistral Small 3.1PROVISOIREMistral AI | 45,037,8 à 52,3 | 25 % · 2 mesures | 0,09 € / 0,26 € | |||
| Sans rang | GPT-4o miniPROVISOIREOpenAI | 44,837,5 à 52,0 | 38 % · 3 mesures | 0,13 € / 0,53 € | |||
| Sans rang | Claude 3.5 HaikuPROVISOIREAnthropic | 44,536,8 à 52,3 | 38 % · 3 mesures | 0,70 € / 3,51 € | |||
| Sans rang | Qwen-TurboPROVISOIREAlibaba | 44,536,7 à 52,4 | 12 % · 1 mesure | 0,04 € / 0,18 € | |||
| Sans rang | Tulu 3 (Tülu 3) 70BPROVISOIREAllen Institute for AI | 44,036,0 à 52,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mistral Small 3PROVISOIREMistral AI | 43,535,8 à 51,2 | 25 % · 2 mesures | 0,04 € / 0,07 € | |||
| Sans rang | Llama 3.1-70BPROVISOIREMeta AI | 43,235,5 à 51,0 | 25 % · 2 mesures | 0,63 € / 0,63 € | |||
| Sans rang | Amazon Nova ProPROVISOIREAmazon | 42,434,3 à 50,6 | 12 % · 1 mesure | 0,70 € / 2,81 € | |||
| Sans rang | Llama 3.2 90BPROVISOIREMeta AI | 42,234,0 à 50,4 | 12 % · 1 mesure | 0,31 € / 0,35 € | |||
| Sans rang | Gemma 3 4BPROVISOIREGoogle DeepMind | 41,933,5 à 50,3 | 25 % · 2 mesures | 0,04 € / 0,09 € | |||
| Sans rang | Llama 3-70BPROVISOIREMeta AI | 41,733,7 à 49,6 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Qwen2-72BPROVISOIREAlibaba | 41,733,5 à 49,8 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemini 1.5 Flash (May 2024)PROVISOIREGoogle DeepMind | 41,433,4 à 49,4 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | GPT-4 (Mar 2023)PROVISOIREOpenAI | 41,432,6 à 50,2 | 25 % · 2 mesures | 26 € / 53 € | |||
| Sans rang | GPT-4 (Jun 2023)PROVISOIREOpenAI | 41,132,7 à 49,6 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Qwen2.5-Coder-32B-InstructPROVISOIREAlibaba | 40,932,7 à 49,0 | 12 % · 1 mesure | 0,25 € / 0,76 € | |||
| Sans rang | Claude 3 SonnetPROVISOIREAnthropic | 40,332,2 à 48,5 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Hermes 2 Theta Llama-3 70BPROVISOIRENous Research | 40,132,2 à 48,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemma 2 27BPROVISOIREGoogle DeepMind | 39,631,4 à 47,8 | 25 % · 2 mesures | 0,57 € / 0,57 € | |||
| Sans rang | Nemotron-4 340BPROVISOIRENVIDIA | 39,631,4 à 47,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Claude 2PROVISOIREAnthropic | 39,331,6 à 47,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | GLM-4 (0520)PROVISOIREZ.ai (Zhipu AI) | 39,030,9 à 47,2 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mistral LargePROVISOIREMistral AI | 37,729,5 à 46,0 | 25 % · 2 mesures | 0,44 € / 1,32 € | |||
| Sans rang | Amazon Nova LitePROVISOIREAmazon | 37,229,1 à 45,3 | 12 % · 1 mesure | 0,05 € / 0,21 € | |||
| Sans rang | Gemini 1.5 Flash 8BPROVISOIREGoogle DeepMind | 37,228,7 à 45,8 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Qwen3-1.7BPROVISOIREAlibaba | 37,230,5 à 43,9 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Claude 2.1PROVISOIREAnthropic | 37,030,4 à 43,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen2.5-7BPROVISOIREAlibaba | 36,730,3 à 43,1 | 12 % · 1 mesure | 0,15 € / 0,61 € | |||
| Sans rang | Claude 3 HaikuPROVISOIREAnthropic | 36,228,0 à 44,4 | 25 % · 2 mesures | 0,22 € / 1,10 € | |||
| Sans rang | Gemma 2 9BPROVISOIREGoogle DeepMind | 34,926,7 à 43,0 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Gemini 1.0 ProPROVISOIREGoogle DeepMind | 34,629,4 à 39,8 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mixtral 8x22BPROVISOIREMistral AI | 34,126,1 à 42,1 | 12 % · 1 mesure | 1,76 € / 5,27 € | |||
| Sans rang | Amazon Nova MicroPROVISOIREAmazon | 33,825,8 à 41,8 | 12 % · 1 mesure | 0,03 € / 0,12 € | |||
| Sans rang | granite-4.0-microPROVISOIREIBM | 33,829,1 à 38,6 | 12 % · 1 mesure | 0,01 € / 0,10 € | |||
| Sans rang | Mistral MediumPROVISOIREMistral AI | 32,824,8 à 40,8 | 12 % · 1 mesure | 1,32 € / 6,58 € | |||
| Sans rang | Ministral 8BPROVISOIREMistral AI | 32,024,0 à 40,0 | 12 % · 1 mesure | 0,09 € / 0,09 € | |||
| Sans rang | Qwen1.5-110BPROVISOIREAlibaba | 32,024,0 à 40,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen1.5-72BPROVISOIREAlibaba | 30,722,8 à 38,6 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Yi-1.5-34BPROVISOIRE01.AI | 30,722,8 à 38,6 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | GPT-3.5 Turbo (Jan 2024)PROVISOIREOpenAI | 29,421,4 à 37,4 | 38 % · 3 mesures | 0,44 € / 1,32 € | |||
| Sans rang | Llama 3.1-8BPROVISOIREMeta AI | 28,120,1 à 36,0 | 25 % · 2 mesures | 0,02 € / 0,03 € | |||
| Sans rang | Llama 3-8BPROVISOIREMeta AI | 27,819,9 à 35,7 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Mixtral 8x7BPROVISOIREMistral AI | 27,619,7 à 35,4 | 12 % · 1 mesure | 0,61 € / 0,61 € | |||
| Sans rang | GPT-3.5 Turbo (Nov 2023)PROVISOIREOpenAI | 26,518,7 à 34,3 | 12 % · 1 mesure | 0,88 € / 1,76 € | |||
| Sans rang | DBRXPROVISOIREDatabricks | 25,718,0 à 33,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen1.5-32BPROVISOIREAlibaba | 25,718,0 à 33,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen1.5-14BPROVISOIREAlibaba | 22,915,2 à 30,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | phi-3-small 7.4BPROVISOIREMicrosoft | 22,114,5 à 29,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | granite-4.0-1bPROVISOIREIBM | 21,820,9 à 22,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | granite-4.0-350mPROVISOIREIBM | 21,820,9 à 22,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | QwQ-32B-PreviewPROVISOIREAlibaba | 20,012,5 à 27,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | DeepSeek LLM 67BPROVISOIREDeepSeek | 19,712,2 à 27,2 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | Yi-34BPROVISOIRE01.AI | 19,211,8 à 26,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Llama 3.2 3BPROVISOIREMeta AI | 17,610,3 à 25,0 | 12 % · 1 mesure | 0,04 € / 0,29 € | |||
| Sans rang | Llama 2-70BPROVISOIREMeta AI | 16,39,0 à 23,6 | 25 % · 2 mesures | Inconnu | |||
| Sans rang | MPT-30BPROVISOIREMosaicML | 12,75,6 à 19,7 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | phi-3-mini 3.8BPROVISOIREMicrosoft | 12,45,4 à 19,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen-14BPROVISOIREAlibaba | 12,45,4 à 19,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Qwen1.5-7BPROVISOIREAlibaba | 12,25,1 à 19,2 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mistral 7B v0.2PROVISOIREMistral AI | 11,14,2 à 18,1 | 12 % · 1 mesure | 0,14 € / 0,19 € | |||
| Sans rang | Gemma 7BPROVISOIREGoogle DeepMind | 10,63,7 à 17,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Mistral 7B v0.3PROVISOIREMistral AI | 10,310,2 à 10,5 | 12 % · 1 mesure | 0,22 € / 0,22 € | |||
| Sans rang | Llama 2-13BPROVISOIREMeta AI | 8,51,7 à 15,3 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Llama 2-7BPROVISOIREMeta AI | 2,20,0 à 8,5 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | chatglm2-6bPROVISOIREZ.ai (Zhipu AI) | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | DeepSeek-R1-Distill-Qwen-1.5BPROVISOIREDeepSeek | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Dolly 2.0-12bPROVISOIREDatabricks | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemma 2BPROVISOIREGoogle DeepMind | 0,00,0 à 3,1 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Gemma 3 1BPROVISOIREGoogle DeepMind | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | Llama 3.2 1BPROVISOIREMeta AI | 0,00,0 à 6,1 | 25 % · 2 mesures | 0,02 € / 0,18 € | |||
| Sans rang | LLaMA-13BPROVISOIREMeta AI | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | MPT-7BPROVISOIREMosaicML | 0,00,0 à 2,0 | 12 % · 1 mesure | Inconnu | |||
| Sans rang | stablelm-tuned-alpha-7bPROVISOIREStability AI | 0,00,0 à 0,0 | 12 % · 1 mesure | Inconnu | |||
Aucun modèle ne correspond à ces filtres.
81 modèles classés et 251 provisoires. La colonne Semaine se remplit à partir de la deuxième édition.
Ce qui compose le score
La démarche →Le score Maths agrège 8 benchmarks, chacun pesé ci-dessous. 3 benchmarks archivés ne comptent plus. La tâche pèse 10 % du score général.
- Actif12,5 %Arena, questions de maths
- Actif12,5 %FrontierMath Erdős
- Saturé12,5 %FrontierMath, palier 4 (v2)
- Actif12,5 %FrontierMath, palier 4 (version 2025)
- Actif12,5 %FrontierMath, paliers 1 à 3 (v2)
- Actif12,5 %FrontierMath, paliers 1 à 3 (version 2025)
- Saturé12,5 %OTIS Mock AIME 2024-2025
- Saturé12,5 %ProofBench
Voir aussi : classement général, modèles ouverts, modèles locaux, rapport qualité-prix et votre classement, selon vos priorités.