Mistral · Mistral Small
Apache-2.0 hybrid MoE (119B total / 6.5B active) that unifies instruct, reasoning and coding in one low-cost model.
Leader: Claude Opus 5.5 at 58
Leader: Claude Opus 5.5 at 66.9%
Leader: Claude Mythos 5.1 at 60.9%
Leader: Claude Fable 5.1 at 91.4%
Leader: GPT-6 Astra at 96.1%
Leader: Claude Opus 5.5 at 61.4%
Leader: GPT-5.6 Sol at 32.3%
Leader: Kimi K3 at 88.7%
Leader: Claude Opus 5.5 at 87.7%
| Result | Score | Reported by | Source |
|---|---|---|---|
| Terminal-Bench Hard (AA) · reasoning | 17.4% | independent | Artificial Analysis: Mistral Small 4 (Reasoning) ↗ |
| IFBench · reasoning | 48.2% | independent | Artificial Analysis: Mistral Small 4 (Reasoning) ↗ |
| tau2-Bench Banking (AA) · reasoning | 4.9% | independent | Artificial Analysis: Mistral Small 4 (Reasoning) ↗ |
| AA-Omniscience Index · reasoning | -30.4 index | independent | Artificial Analysis: Mistral Small 4 (Reasoning) ↗ |
| AA Analyst Agent · reasoning | 1.2% | independent | Artificial Analysis: Mistral Small 4 (Reasoning) ↗ |