Skip to contentT/TTokens or Towers

Compare

LeaderboardEvery tracked model, ranked and priced side by sideBenchmarksWhat each test measures and who leads itPricesAPI, media, and hardware rates with a cost calculator

Plan

Stack plannerSix questions to a priced local and hosted stackShortlistOur picks for subscriptions, APIs, and hardwareEnterpriseSeat plans against self-hosting for a whole team

Build

HardwareGPUs and workstations, and which models they fitHostsWhere to rent open-weight inference, and for how muchHarnessesCoding agents and the models they drive

Lab

FindingsResults from our own evaluationsScanAudit a repo’s AI usage in the browserMethodWhere the numbers come from and how we check them
Data checked September 25, 2026
Data checked September 25, 2026
Plan a stack
T/TTokens or Towers

An independent guide to AI models, what they cost, and the hardware to run them. No account, no API key, no sponsors.

Compare

  • Leaderboard
  • Benchmarks
  • Prices

Plan

  • Stack planner
  • Shortlist
  • Enterprise

Build

  • Hardware
  • Hosts
  • Harnesses

Lab

  • Findings
  • Scan
  • Method
Data checked September 25, 2026 · Every number links to its sourceRSS · Source on GitHub
Leaderboard/Baidu
Baidu

Baidu · ERNIE 5

ERNIE 5.0 Thinking Preview

Closed weightsSupersededproprietary

Late-2025 reasoning preview of ERNIE 5.0; the only ERNIE 5.x model with an Artificial Analysis page.

AA Intelligence
14.3#77 of 142
Price
—
Output speed
—
Context window
128K tokens

Composite indexes

  • Artificial Analysis Intelligence Index#77 of 142
    14independentv4.3.2 (AA estimate; independent evaluation forthcoming)Artificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: Claude Opus 5.5 at 58

Coding

  • LiveCodeBench#8 of 56
    81.2%independentAAArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: Fugu at 92.9%

Agents and tool use

  • τ²-Bench#28 of 88
    83.9%independentAA, telecomArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: GLM-5.2 at 99.1%

Reasoning and knowledge

  • Humanity’s Last Exam#84 of 147
    13.3%independentAA, no toolsArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: Claude Opus 5.5 at 61.4%

  • GPQA Diamond#78 of 142
    77.7%independentAAArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: GPT-6 Astra at 96.1%

  • CritPt#68 of 139
    1.4%independentAAArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: GPT-5.6 Sol at 32.3%

Vision and long context

  • MMMU-Pro#50 of 72
    64.6%independentAAArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

    Leader: Claude Opus 5.5 at 87.7%

Other published results

ResultScoreReported bySource
aime-2025 · AA, AIME 202585%independentArtificial Analysis: ERNIE 5.0 Thinking Preview ↗
Terminal-Bench Hard (AA) · AA25%independentArtificial Analysis: ERNIE 5.0 Thinking Preview ↗
IFBench (AA) · AA41.4%independentArtificial Analysis: ERNIE 5.0 Thinking Preview ↗
AA-Omniscience Index · AA-45.6 index (-100..100)independentArtificial Analysis: ERNIE 5.0 Thinking Preview ↗

Specs

Released
November 13, 2025
Parameters
Undisclosed
Size class
Undisclosed
Input
text, image, video
Output
text
Context
128K tokens
Max output
Not published

    Will it fit your budget?

    The planner prices this model against local hardware and other APIs for your workload.

    Plan a stack →

    Compare with

    • GoogleGemma 4 12B14.2
    • MistralMistral Medium 3.514.2
    • AmazonAmazon Nova 2 Pro (Preview)14.2
    • IBMGranite 4.2 30B14.8