Skip to contentT/TTokens or Towers

Compare

LeaderboardEvery tracked model, ranked and priced side by sideBenchmarksWhat each test measures and who leads itPricesAPI, media, and hardware rates with a cost calculator

Plan

Stack plannerSix questions to a priced local and hosted stackShortlistOur picks for subscriptions, APIs, and hardwareEnterpriseSeat plans against self-hosting for a whole team

Build

HardwareGPUs and workstations, and which models they fitHostsWhere to rent open-weight inference, and for how muchHarnessesCoding agents and the models they drive

Lab

FindingsResults from our own evaluationsScanAudit a repo’s AI usage in the browserMethodWhere the numbers come from and how we check them
Data checked September 25, 2026
Data checked September 25, 2026
Plan a stack
T/TTokens or Towers

An independent guide to AI models, what they cost, and the hardware to run them. No account, no API key, no sponsors.

Compare

  • Leaderboard
  • Benchmarks
  • Prices

Plan

  • Stack planner
  • Shortlist
  • Enterprise

Build

  • Hardware
  • Hosts
  • Harnesses

Lab

  • Findings
  • Scan
  • Method
Data checked September 25, 2026 · Every number links to its sourceRSS · Source on GitHub
Leaderboard/Google
Google

Google · Gemini 3

Gemini 3.5 Flash

Closed weightsSupersededproprietary

Legacy Gemini 3.5 Flash (May 2026), priced above the newer 3.6-3.8 Flash models, which also score higher.

AA Intelligence
32.6#36 of 142
Input / output per 1M
$1.50/ $9.00
Output speed
238tok/s
Context window
1M tokens

Composite indexes

  • Artificial Analysis Intelligence Index#36 of 142
    33independentv4.3.2, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Opus 5.5 at 58

Coding

  • SciCode#27 of 94
    53.9%independentAA run, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Opus 5.5 at 66.9%

  • Terminal-Bench 4.0#35 of 88
    6.6%independentAA run (Terminal-Bench 4.0), high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Mythos 5.1 at 60.9%

  • Terminal-Bench 2.x#25 of 103
    78.7%independentAA run (Terminal-Bench 2.1), high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Fable 5.1 at 91.4%

Agents and tool use

  • GDPval-AA#40 of 103
    1185independentGDPval-AA v2.1, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Opus 5.5 at 1846

  • τ²-Bench#8 of 88
    95.3%independentAA run, telecom, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: GLM-5.2 at 99.1%

Reasoning and knowledge

  • GPQA Diamond#25 of 142
    92.2%independentAA run, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: GPT-6 Astra at 96.1%

  • Humanity’s Last Exam#26 of 147
    42.7%independentAA run, no tools, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Opus 5.5 at 61.4%

  • CritPt#35 of 139
    13.1%independentAA run, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: GPT-5.6 Sol at 32.3%

  • ARC-AGI-2#12 of 25
    72.1%independenthigh thinking; ARC Prize verifiedARC Prize leaderboard data (semi-private) ↗

    Leader: GPT-6 Astra at 95%

Vision and long context

  • AA Long Context Reasoning#63 of 139
    73.3%independentAA-LCR v1.1, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Kimi K3 at 88.7%

  • MMMU-Pro#6 of 72
    84.3%independentAA run, high thinkingArtificial Analysis: Gemini 3.5 Flash ↗

    Leader: Claude Opus 5.5 at 87.7%

Human preference

  • LMArena Text#17 of 54
    1478independentgemini-3.5-flash-high; 38,257 votesLMArena text leaderboard ↗

    Leader: Claude Fable 5 at 1506

  • LMArena WebDev#38 of 54
    1500independentgemini-3.5-flash-high; 8,909 votesLMArena Code Arena | WebDev leaderboard ↗

    Leader: Claude Opus 5.5 at 1818

Specs

Released
May 19, 2026
Parameters
Undisclosed
Size class
Undisclosed
Input
text, image, audio, video
Output
text
Context
1M tokens
Max output
Not published
Cached input
$0.15 / 1M
Blended 70/30
$3.75 / 1M
First token
16.98s
  • Price: Gemini Developer API pricing ↗ Now pricier than 3.6/3.7/3.8 Flash; Google labels it a legacy/earlier Flash model.
  • Speed: Artificial Analysis: Gemini 3.5 Flash (high thinking) ↗

Will it fit your budget?

The planner prices this model against local hardware and other APIs for your workload.

Plan a stack →

Compare with

  • OpenAIGPT-5.3 Codex32.5
  • MetaMuse Spark 1.133.7
  • GoogleGemini 3.6 Flash34
  • DeepSeekDeepSeek V4 Flash 073134