Skip to contentT/TTokens or Towers

Compare

LeaderboardEvery tracked model, ranked and priced side by sideBenchmarksWhat each test measures and who leads itPricesAPI, media, and hardware rates with a cost calculator

Plan

Stack plannerSix questions to a priced local and hosted stackShortlistOur picks for subscriptions, APIs, and hardwareEnterpriseSeat plans against self-hosting for a whole team

Build

HardwareGPUs and workstations, and which models they fitHostsWhere to rent open-weight inference, and for how muchHarnessesCoding agents and the models they drive

Lab

FindingsResults from our own evaluationsScanAudit a repo’s AI usage in the browserMethodWhere the numbers come from and how we check them
Data checked September 25, 2026
Data checked September 25, 2026
Plan a stack
T/TTokens or Towers

An independent guide to AI models, what they cost, and the hardware to run them. No account, no API key, no sponsors.

Compare

  • Leaderboard
  • Benchmarks
  • Prices

Plan

  • Stack planner
  • Shortlist
  • Enterprise

Build

  • Hardware
  • Hosts
  • Harnesses

Lab

  • Findings
  • Scan
  • Method
Data checked September 25, 2026 · Every number links to its sourceRSS · Source on GitHub
Leaderboard/IBM
IBM

IBM · Granite 4.2

Granite 4.2 8B

Open weightsGenerally availableNewApache-2.0

Mid-size Granite 4.2 dense open-weights reasoning model for self-hosted enterprise tasks.

AA Intelligence
11.1#95 of 142
Input / output per 1M
$0.06/ $0.25
Output speed
79tok/s
Context window
131K tokens

Composite indexes

  • Artificial Analysis Intelligence Index#95 of 142
    11independentIntelligence Index v4.3; reasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Opus 5.5 at 58

Coding

  • SciCode#87 of 94
    36.1%vendorthinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗
    31.5%independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Opus 5.5 at 66.9%

  • Terminal-Bench 4.0#84 of 88
    0%independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Mythos 5.1 at 60.9%

  • Terminal-Bench 2.x#85 of 103
    20.6%vendorTerminal-Bench 2.1; thinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗
    18.4%independentTerminal-Bench 2.1; reasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Fable 5.1 at 91.4%

  • SWE-bench Pro#31 of 31
    19.1%vendorthinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗

    Leader: Claude Opus 5.5 at 89.9%

  • SWE-bench Verified#14 of 16
    47.7%vendorthinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗

    Leader: Claude Sonnet 5 at 85.2%

  • LiveCodeBench#19 of 56
    73.2%vendorv6; thinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗

    Leader: Fugu at 92.9%

Agents and tool use

  • GDPval-AA#77 of 103
    1189vendorGDPval (vendor-run)Hugging Face: ibm-granite/granite-4.2-30b model card ↗
    478independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Opus 5.5 at 1846

Reasoning and knowledge

  • GPQA Diamond#109 of 142
    64.1%vendorGPQA; thinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗
    63.1%independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: GPT-6 Astra at 96.1%

  • Humanity’s Last Exam#107 of 147
    9.7%independentno tools; reasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Claude Opus 5.5 at 61.4%

  • CritPt#86 of 139
    0.3%independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: GPT-5.6 Sol at 32.3%

  • HMMT February 2026#7 of 8
    78.3%vendorHMMT Feb 2025 (not 2026); thinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗

    Leader: GLM-5.2 at 92.5%

  • MMLU-Pro#11 of 14
    74.0%vendorthinkingHugging Face: ibm-granite/granite-4.2-30b model card ↗

    Leader: Nemotron 3 Ultra 550B A55B at 86.8%

Vision and long context

  • AA Long Context Reasoning#94 of 139
    45%independentreasoningArtificial Analysis: Granite 4.2 8B ↗

    Leader: Kimi K3 at 88.7%

Other published results

ResultScoreReported bySource
aime-2025 · AIME 2025; thinking86.67%vendorHugging Face: ibm-granite/granite-4.2-30b model card ↗
tau2-Bench Banking (AA) · reasoning7.6%independentArtificial Analysis: Granite 4.2 8B ↗
AA-Omniscience Index · reasoning-17.2 indexindependentArtificial Analysis: Granite 4.2 8B ↗
SWE-bench Multilingual · thinking30.78%vendorHugging Face: ibm-granite/granite-4.2-30b model card ↗
tau3-Bench · thinking66.34%vendorHugging Face: ibm-granite/granite-4.2-30b model card ↗
IFBench · prompt-level; thinking79.33%vendorHugging Face: ibm-granite/granite-4.2-30b model card ↗

Specs

Released
August 25, 2026
Parameters
8B
Size class
Under 10B
Input
text
Output
text
Context
131K tokens
Max output
Not published
Cached input
$0.015 / 1M
Blended 70/30
$0.117 / 1M
First token
0.74s
  • Price: Artificial Analysis: Granite 4.2 8B ↗ First-party (IBM) price as listed by Artificial Analysis. IBM watsonx.ai rate not independently confirmed on an IBM page.
  • Speed: Artificial Analysis leaderboard (Granite 4.2 8B, median across providers or first-party as shown) ↗

Will it fit your budget?

The planner prices this model against local hardware and other APIs for your workload.

Plan a stack →

Compare with

  • MistralMistral Small 411.3
  • Arcee AITrinity Large Thinking10.8
  • MeituanLongCat Flash Lite11.5
  • OpenAIgpt-oss-120b11.6