Alibaba · Qwen3.7
Mid-tier hosted multimodal Qwen and still the newest Plus model. Parameter counts (397B/17B) come from the Qwen3.8-Flash-Next comparison table.
Leader: Claude Opus 5.5 at 58
Leader: Claude Mythos 5.1 at 60.9%
Leader: Claude Fable 5.1 at 91.4%
Leader: Claude Opus 5.5 at 66.9%
Leader: Claude Opus 5.5 at 89.9%
Leader: GLM-5.2 at 99.1%
Leader: Claude Opus 5.5 at 1846
Leader: Kimi K3 at 84.8%
Leader: GPT-6 Astra at 96.1%
Leader: Claude Opus 5.5 at 61.4%
Leader: GPT-5.6 Sol at 32.3%
Leader: Kimi K3 at 88.7%
Leader: Claude Opus 5.5 at 87.7%
Leader: Claude Fable 5 at 1506
| Result | Score | Reported by | Source |
|---|---|---|---|
| IFBench (AA) · reasoning | 78% | independent | Artificial Analysis: Qwen3.7 Plus ↗ |
| Terminal-Bench Hard (AA) · reasoning | 47% | independent | Artificial Analysis: Qwen3.7 Plus ↗ |