Skip to contentT/TTokens or Towers

Compare

LeaderboardEvery tracked model, ranked and priced side by sideBenchmarksWhat each test measures and who leads itPricesAPI, media, and hardware rates with a cost calculator

Plan

Stack plannerSix questions to a priced local and hosted stackShortlistOur picks for subscriptions, APIs, and hardwareEnterpriseSeat plans against self-hosting for a whole team

Build

HardwareGPUs and workstations, and which models they fitHostsWhere to rent open-weight inference, and for how muchHarnessesCoding agents and the models they drive

Lab

FindingsResults from our own evaluationsScanAudit a repo’s AI usage in the browserMethodWhere the numbers come from and how we check them
Data checked September 25, 2026
Data checked September 25, 2026
Plan a stack
T/TTokens or Towers

An independent guide to AI models, what they cost, and the hardware to run them. No account, no API key, no sponsors.

Compare

  • Leaderboard
  • Benchmarks
  • Prices

Plan

  • Stack planner
  • Shortlist
  • Enterprise

Build

  • Hardware
  • Hosts
  • Harnesses

Lab

  • Findings
  • Scan
  • Method
Data checked September 25, 2026 · Every number links to its sourceRSS · Source on GitHub
Leaderboard/Kuaishou
Kuaishou

Kuaishou · KAT-Coder V2

KAT-Coder-Pro V2

Closed weightsSupersededproprietary

Previous Kwaipilot agentic coding model; still listed on Artificial Analysis but superseded by KAT-Coder-Pro V2.5.

AA Intelligence
21.7#63 of 142
Price
—
Output speed
—
Context window
256K tokens

Composite indexes

  • Artificial Analysis Intelligence Index#63 of 142
    22independentv4.3.2 (AA estimate; independent evaluation forthcoming), non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: Claude Opus 5.5 at 58

Coding

  • Terminal-Bench 2.x#35 of 103
    70%independentAA, Terminal-Bench 2.1, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: Claude Fable 5.1 at 91.4%

Agents and tool use

  • τ²-Bench#21 of 88
    89.5%independentAA, telecom, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: GLM-5.2 at 99.1%

  • GDPval-AA#68 of 103
    720independentGDPval-AA v2.1, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: Claude Opus 5.5 at 1846

Reasoning and knowledge

  • Humanity’s Last Exam#78 of 147
    16.1%independentAA, no tools, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: Claude Opus 5.5 at 61.4%

  • GPQA Diamond#57 of 142
    85.5%independentAA, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: GPT-6 Astra at 96.1%

  • CritPt#102 of 139
    0%independentAA, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: GPT-5.6 Sol at 32.3%

Vision and long context

  • AA Long Context Reasoning#67 of 139
    73%independentAA, non-reasoningArtificial Analysis: KAT Coder Pro V2 ↗

    Leader: Kimi K3 at 88.7%

Other published results

ResultScoreReported bySource
Terminal-Bench Hard (AA) · AA49.2%independentArtificial Analysis: KAT Coder Pro V2 ↗
IFBench (AA) · AA66.7%independentArtificial Analysis: KAT Coder Pro V2 ↗
AA-Omniscience Index · AA-22.5 index (-100..100)independentArtificial Analysis: KAT Coder Pro V2 ↗

Specs

Released
March 27, 2026
Parameters
Undisclosed
Size class
Undisclosed
Input
text
Output
text
Context
256K tokens
Max output
Not published

    Will it fit your budget?

    The planner prices this model against local hardware and other APIs for your workload.

    Plan a stack →

    Compare with

    • GoogleGemini 3.5 Flash-Lite22.2
    • AlibabaQwen3.6-27B21
    • OpenAIGPT-5.4 nano20.7
    • NVIDIANemotron 3 Ultra 550B A55B22.9