Reference prices
Every row cites the provider or product page it came from, with the date we last checked it. Street estimates and planning allowances carry a leading ~ so you can tell them apart from list prices. Confirm against the live page before you spend anything.
Prices last checked August 5, 2026
Open the planner36 models
Input and output rates are USD per million tokens unless a row note says otherwise.
| Provider | Model | Input $/M | Output $/M | Context | Source | Checked |
|---|---|---|---|---|---|---|
| DeepSeek | DeepSeek V4 Flash 0731Cache-miss (uncached) rates; cache hits are far cheaper. | $0.14 | $0.28 | 1M | DeepSeek API pricing ↗ | 2026-08-04 |
| DeepSeek | DeepSeek V4 ProCache-miss rates from the DeepSeek pricing table. | $0.44 | $0.87 | 1M | DeepSeek API pricing ↗ | 2026-08-05 |
| OpenAI | GPT-5.6 LunaShort-context standard rates; prompts over 272K are surcharged. | $0.20 | $1.20 | 1M | OpenAI API pricing ↗ | 2026-08-04 |
| OpenAI | GPT-5.6 TerraShort-context standard rates from the OpenAI pricing table. | $2.00 | $12.00 | 1M | OpenAI API pricing ↗ | 2026-08-05 |
| OpenAI | GPT-5.6 SolEscalation tier; short-context standard rates. | $5.00 | $30.00 | 1M | OpenAI API pricing ↗ | 2026-08-04 |
| OpenAI | GPT-5.4 miniStandard rates from the OpenAI pricing table. | $0.75 | $4.50 | 1M | OpenAI API pricing ↗ | 2026-08-05 |
| OpenAI | GPT-5.4 nanoLowest GPT-5.4 text tier on the OpenAI pricing page. | $0.20 | $1.25 | 1M | OpenAI API pricing ↗ | 2026-08-05 |
| Anthropic | Claude Haiku 4.5 | $1.00 | $5.00 | 200K | Claude API pricing ↗ | 2026-08-05 |
| Anthropic | Claude Sonnet 5Introductory $2/$10 through August 31, 2026; then $3/$15. | $2.00 | $10.00 | 1M | Claude API pricing ↗ | 2026-08-05 |
| Anthropic | Claude Opus 5 | $5.00 | $25.00 | 1M | Claude API pricing ↗ | 2026-08-05 |
| Gemini 3.5 Flash-LitePaid-tier standard rates (text/image/video/audio input). | $0.30 | $2.50 | 1M | Gemini Developer API pricing ↗ | 2026-08-05 | |
| Gemini 3.6 FlashPaid-tier standard rates; output includes thinking tokens. | $1.50 | $7.50 | 1M | Gemini Developer API pricing ↗ | 2026-08-05 | |
| Gemini 3.1 ProPaid tier for prompts ≤200K tokens; longer prompts are $4/$18. | $2.00 | $12.00 | 1M | Gemini Developer API pricing ↗ | 2026-08-05 | |
| xAI | Grok 4.3Standard rates for prompts up to 200K tokens. | $1.25 | $2.50 | 1M | xAI API pricing ↗ | 2026-08-05 |
| xAI | Grok 4.5Standard rates for prompts up to 200K tokens. | $2.00 | $6.00 | 500K | xAI API pricing ↗ | 2026-08-05 |
| Mistral | Mistral Large 3Standard La Plateforme API rates. | $0.50 | $1.50 | 256K | Mistral API pricing ↗ | 2026-08-05 |
| Mistral | Mistral Medium 3.5Standard La Plateforme API rates. | $1.50 | $7.50 | 256K | Mistral API pricing ↗ | 2026-08-05 |
| Mistral | Mistral Small 4Standard La Plateforme API rates. | $0.15 | $0.60 | 256K | Mistral API pricing ↗ | 2026-08-05 |
| Moonshot AI | Kimi K3Cache-miss rates from the official Kimi API table. | $3.00 | $15.00 | 1M | Kimi API pricing ↗ | 2026-08-05 |
| Alibaba | Qwen 3.7 MaxSingapore international list rates across the full context window. | $2.50 | $7.50 | 1M | Alibaba Cloud Model Studio pricing ↗ | 2026-08-05 |
| Alibaba | Qwen 3.7 PlusSingapore international list rates for prompts up to 256K tokens. | $0.40 | $1.60 | 1M | Alibaba Cloud Model Studio pricing ↗ | 2026-08-05 |
| Alibaba | Qwen 3.6 FlashSingapore international rates for prompts up to 256K tokens. | $0.25 | $1.50 | 1M | Alibaba Cloud Model Studio pricing ↗ | 2026-08-05 |
| Z.ai | GLM-5Standard cache-miss rates from the Z.ai pricing table. | $1.00 | $3.20 | 200K | Z.ai model pricing ↗ | 2026-08-05 |
| Z.ai | GLM-5.2Standard cache-miss rates from the Z.ai pricing table. | $1.40 | $4.40 | 200K | Z.ai model pricing ↗ | 2026-08-05 |
| Z.ai | GLM-4.5-AirStandard cache-miss rates from the Z.ai pricing table. | $0.20 | $1.10 | 128K | Z.ai model pricing ↗ | 2026-08-05 |
| Amazon | Amazon Nova ProBedrock standard on-demand rates in primary US regions. | $0.80 | $3.20 | 300K | Amazon Bedrock pricing ↗ | 2026-08-05 |
| Amazon | Amazon Nova LiteBedrock standard on-demand rates in primary US regions. | $0.06 | $0.24 | 300K | Amazon Bedrock pricing ↗ | 2026-08-05 |
| Cohere | Command AStandard production API rates. | $2.50 | $10.00 | 256K | Cohere Command A model card ↗ | 2026-08-05 |
| Meta | Llama 4 MaverickGroqCloud list rates for the hosted Meta model. | $0.50 | $0.77 | 1M | Groq Llama 4 pricing ↗ | 2026-08-05 |
| Meta | Llama 4 ScoutGroqCloud list rates for the hosted Meta model. | $0.11 | $0.34 | 10M | Groq Llama 4 pricing ↗ | 2026-08-05 |
| MiniMax | MiniMax M2.5Bedrock standard rates in primary US regions. | $0.30 | $1.20 | 196K | Amazon Bedrock pricing ↗ | 2026-08-05 |
| Baidu | ERNIE 5.0International Qianfan pay-as-you-go rates. | $1.40 | $5.60 | 128K | Baidu AI Cloud international pricing ↗ | 2026-08-05 |
| Perplexity | SonarSonar API token rates exclude the separate search-context request fee. | $1.00 | $1.00 | 128K | Perplexity Sonar API pricing ↗ | 2026-08-05 |
| AI21 Labs | Jamba 1.5 LargeBedrock standard on-demand rates in US East. | $2.00 | $8.00 | 256K | Amazon Bedrock pricing ↗ | 2026-08-05 |
| AI21 Labs | Jamba 1.5 MiniBedrock standard on-demand rates in US East. | $0.20 | $0.40 | 256K | Amazon Bedrock pricing ↗ | 2026-08-05 |
| Anthropic | Claude Opus 4.8Standard global Claude API rates. | $5.00 | $25.00 | 1M | Anthropic API pricing ↗ | 2026-08-05 |
17 components
A leading ~ marks street estimates and planning allowances. Use those for budget envelopes, and check a live listing when you reach checkout.
| Category | Item | Price | Key spec | Source | Checked |
|---|---|---|---|---|---|
| System | 24 GB value workstation (Used RTX 3090 class)Whole-system planning allowance, not a live used-market quote. | ~$1,500 | 24 GB VRAM | Planning assumption ↗ | 2026-08-04 |
| GPU | GeForce RTX 4090 Founders EditionNVIDIA list starting price for the card alone. | $1,599 | 24 GB GDDR6X | NVIDIA RTX 4090 product page ↗ | 2026-08-05 |
| GPU | GeForce RTX 5090 Founders EditionNVIDIA list starting price for the card alone. | $1,999 | 32 GB GDDR7 · 1,792 GB/s | NVIDIA RTX 5090 product page ↗ | 2026-08-05 |
| System | 32 GB performance workstation (RTX 5090 class build)Whole-system planning allowance around a 5090-class build. | ~$3,200 | 32 GB VRAM · 1,792 GB/s | NVIDIA RTX 5090 specifications ↗ | 2026-08-04 |
| System | DGX Spark Founders EditionNVIDIA marketplace list price for the Founders Edition bundle. | $4,699 | 128 GB unified · 273 GB/s | NVIDIA DGX Spark marketplace ↗ | 2026-08-05 |
| System | Capacity + speed local lab (DGX Spark + RTX 5090)Combined planning allowance, not a bundled SKU. | ~$7,900 | 128 GB unified + 32 GB VRAM | NVIDIA DGX Spark hardware ↗ | 2026-08-04 |
| System | Mac mini (M4 Pro, 24 GB)Apple U.S. starting price for the M4 Pro configuration. | $1,399 | 24 GB unified · 273 GB/s | Apple Mac mini newsroom ↗ | 2026-08-05 |
| System | Mac Studio (M4 Max, 36 GB)Apple U.S. starting price for the base M4 Max Studio. | $1,999 | 36 GB unified · 410 GB/s | Apple Mac Studio newsroom ↗ | 2026-08-05 |
| System | Mac Studio (M3 Ultra, 96 GB)Apple U.S. starting price for the M3 Ultra configuration. | $3,999 | 96 GB unified · 819 GB/s | Apple Mac Studio newsroom ↗ | 2026-08-05 |
| GPU | Used GeForce RTX 4090 Founders EditionRounded snapshot of recent used listings with seller and condition variance. | ~$2,200 | 24 GB GDDR6X | eBay used RTX 4090 listings ↗ | 2026-08-05 |
| GPU | GeForce RTX 5080 Founders EditionNVIDIA launch list price for the card alone. | $999 | 16 GB GDDR7 · 960 GB/s | NVIDIA RTX 50 Series announcement ↗ | 2026-08-05 |
| GPU | GeForce RTX 5070 TiNVIDIA launch list price for the card alone. | $749 | 16 GB GDDR7 · 896 GB/s | NVIDIA RTX 50 Series announcement ↗ | 2026-08-05 |
| GPU | NVIDIA RTX PRO 6000 Blackwell Workstation EditionNVIDIA Marketplace list price for the workstation card. | $13,250 | 96 GB GDDR7 ECC · 1,792 GB/s | NVIDIA RTX PRO 6000 marketplace ↗ | 2026-08-05 |
| System | Framework Desktop Ryzen AI Max+ 395 128 GBFramework launch list price for the 128 GB DIY Edition. | $1,999 | 128 GB LPDDR5X · 256 GB/s | Framework Desktop announcement ↗ | 2026-08-05 |
| System | Minisforum MS-S1 MAX 128 GBMinisforum regular list price for the 128 GB and 2 TB configuration. | $4,549 | 128 GB LPDDR5X · 256 GB/s | Minisforum MS-S1 MAX product page ↗ | 2026-08-05 |
| System | 16-inch MacBook Pro M4 Max 48 GBApple U.S. launch list price for the 16-core CPU and 40-core GPU configuration. | $3,999 | 48 GB unified memory · 546 GB/s | Apple 16-inch MacBook Pro technical specifications ↗ | 2026-08-05 |
| System | NVIDIA Jetson AGX Thor Developer KitCurrent NVIDIA developer kit MSRP. | $5,499 | 128 GB LPDDR5X · 273 GB/s | NVIDIA Jetson product pricing FAQ ↗ | 2026-08-05 |