Skip to contentT/TTokens or Towers

Compare

LeaderboardEvery tracked model, ranked and priced side by sideBenchmarksWhat each test measures and who leads itPricesAPI, media, and hardware rates with a cost calculator

Plan

Stack plannerSix questions to a priced local and hosted stackShortlistOur picks for subscriptions, APIs, and hardwareEnterpriseSeat plans against self-hosting for a whole team

Build

HardwareGPUs and workstations, and which models they fitHostsWhere to rent open-weight inference, and for how muchHarnessesCoding agents and the models they drive

Lab

FindingsResults from our own evaluationsScanAudit a repo’s AI usage in the browserMethodWhere the numbers come from and how we check them
Data checked September 25, 2026
Data checked September 25, 2026
Plan a stack
T/TTokens or Towers

An independent guide to AI models, what they cost, and the hardware to run them. No account, no API key, no sponsors.

Compare

  • Leaderboard
  • Benchmarks
  • Prices

Plan

  • Stack planner
  • Shortlist
  • Enterprise

Build

  • Hardware
  • Hosts
  • Harnesses

Lab

  • Findings
  • Scan
  • Method
Data checked September 25, 2026 · Every number links to its sourceRSS · Source on GitHub
Leaderboard/NVIDIA
NVIDIA

NVIDIA · Nemotron 3

Nemotron 3 Nano Omni 30B A3B Reasoning

Open weightsGenerally availableNVIDIA Open Model License

Open omni-modal (text, image, audio, video in; text out) 30B/3B-active reasoning model for multimodal agents.

AA Intelligence
10.3#98 of 142
Input / output per 1M
$0.195/ $1.095
Output speed
267tok/s
Context window
256K tokens

Composite indexes

  • Artificial Analysis Intelligence Index#98 of 142
    10independentIntelligence Index v4.3; reasoning; estimatedArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Claude Opus 5.5 at 58

Coding

  • Terminal-Bench 2.x#94 of 103
    6.7%independentTerminal-Bench 2.1; reasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Claude Fable 5.1 at 91.4%

Agents and tool use

  • τ²-Bench#48 of 88
    45.3%independenttelecom; reasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: GLM-5.2 at 99.1%

  • GDPval-AA#92 of 103
    216independentreasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Claude Opus 5.5 at 1846

Reasoning and knowledge

  • GPQA Diamond#130 of 142
    46.9%independentreasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: GPT-6 Astra at 96.1%

  • Humanity’s Last Exam#127 of 147
    4.8%independentno tools; reasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Claude Opus 5.5 at 61.4%

  • CritPt#111 of 139
    0%independentreasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: GPT-5.6 Sol at 32.3%

Vision and long context

  • AA Long Context Reasoning#95 of 139
    39.7%independentreasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Kimi K3 at 88.7%

  • MMMU-Pro#62 of 72
    53.2%independentreasoningArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

    Leader: Claude Opus 5.5 at 87.7%

Other published results

ResultScoreReported bySource
Terminal-Bench Hard (AA) · reasoning8.3%independentArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗
IFBench · reasoning63.2%independentArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗
AA-Omniscience Index · reasoning-57.4 indexindependentArtificial Analysis: Nemotron 3 Nano Omni 30B A3B Reasoning ↗

Specs

Released
April 29, 2026
Parameters
30B total, 3B active
Size class
10 to 40B
Input
text, image, audio, video
Output
text
Context
256K tokens
Max output
Not published
Cached input
$0.195 / 1M
Blended 70/30
$0.465 / 1M
First token
0.47s
  • Price: Artificial Analysis model page ↗ No first-party paid NVIDIA API (build.nvidia.com is a free trial endpoint). Value is the Artificial Analysis median across third-party API providers.
  • Speed: Artificial Analysis leaderboard (Nemotron 3 Nano Omni 30B A3B Reasoning, median across providers or first-party as shown) ↗

Will it fit your budget?

The planner prices this model against local hardware and other APIs for your workload.

Plan a stack →

Compare with

  • Prime IntellectINTELLECT-310.6
  • MetaLlama 4 Maverick10
  • CohereNorth Mini Code 1.09.9
  • Arcee AITrinity Large Thinking10.8