Skip to content
All systems operational0 AI providers monitored, polled every 2 minutes
Live status

AI API Cost Calculator

Estimate your monthly AI API spend across all major providers. Adjust your token volume and input/output ratio to see real-time cost comparisons.

Configure Your Usage

Input: 70%Output: 30%

700K input tokens + 300K output tokens per month

For general tasks, mid-tier models like Claude Sonnet 5, GPT-6 Sol, or Gemini 3.8 Flash offer a strong balance of quality and cost.

Estimated Monthly Costsfor 1M tokens/month

ProviderModelReleasedMonthly Cost
AnthropicClaude Opus 5.5Most CapableSep 2026$8.80
AnthropicClaude Fable 5.1Most CapableSep 2026$22.00
AnthropicClaude Mythos 5.1Sep 2026$22.00
OpenAIGPT-6 AstraMost CapableSep 2026$22.00
OpenAIGPT-6 SolSep 2026$4.40
OpenAIGPT-6 LunaSep 2026$0.220
GoogleGemini 3.8 FlashSep 2026$1.65
MetaMuse Spark 1.3Sep 2026$2.15
MetaMuse Spark 1.3 ContributorSep 2026$0.130
DeepSeekDeepSeek V4.1 FlashOpen SourceSep 2026$0.570
AlibabaQwen3.8-Omni-FlashSep 2026$0.246
xAIGrok 4.7Sep 2026$3.20
Z.ai (Zhipu AI)GLM-5.3-FlashXSep 2026$0.634
Sakana AIFugu Ultra v2.0Sep 2026$12.50
Sakana AIFugu MaxSep 2026$3.20
UnbiasedParetoSep 2026$4.00
XiaomiMiMo-V2.6-ProOpen SourceSep 2026$0.566
XiaomiMiMo-V2.6-FlashOpen SourceSep 2026$0.182
GoogleGemini 3.7 FlashAug 2026$1.65
MetaMuse Glimmer 30BOpen SourceAug 2026$0.695
MetaMuse Spark 1.2Aug 2026$2.15
MetaMuse Spark 1.2 ContributorAug 2026$0.130
AlibabaQwen3.8 27BOpen SourceAug 2026$1.25
AlibabaQwen3.8 2.4T-A95BOpen SourceAug 2026$3.20
AlibabaQwen3.8-MaxAug 2026$3.20
AlibabaQwen3.8-FlashAug 2026$0.246
xAIGrok 4.6Aug 2026$3.20
NVIDIANemotron 3.5 LightningBest ValueOpen SourceAug 2026$0.116
MicrosoftMAI-Code-1.1-FlashAug 2026$0.500
Z.ai (Zhipu AI)GLM-5.3Open SourceAug 2026$2.30
Z.ai (Zhipu AI)GLM-5.3 FlashOpen SourceAug 2026$0.255
TencentHy4 previewOpen SourceAug 2026$1.33
AnthropicClaude Opus 5Most CapableJul 2026$11.00
OpenAIGPT-5.6 SolMost CapableJul 2026$8.80
OpenAIGPT-5.6 TerraJul 2026$5.00
OpenAIGPT-5.6 LunaJul 2026$0.500
GoogleGemini 3.6 FlashJul 2026$1.65
GoogleGemini 3.5 Flash-LiteJul 2026$0.960
MetaMuse Spark 1.1Jul 2026$2.15
xAIGrok 4.5Jul 2026$3.20
Moonshot AIKimi K3Open SourceJul 2026$6.60
poolsideLaguna S 2.1Open SourceJul 2026$0.130
AnthropicClaude Fable 5Jun 2026$22.00
AnthropicClaude Sonnet 5Jun 2026$4.40
MiniMaxMiniMax M3Open SourceJun 2026$0.570
MeituanLongCat-2.0Open SourceJun 2026$1.41
Z.ai (Zhipu AI)GLM-5.2Open SourceJun 2026$2.30
AnthropicClaude Opus 4.8May 2026$11.00
GoogleGemini 3.1 Flash-LiteMay 2026$0.625
GoogleGemini 3.5 FlashMay 2026$3.75
AlibabaQwen3.7-MaxMay 2026$4.00
AnthropicClaude Opus 4.7Apr 2026$11.00
OpenAIGPT-5.5Apr 2026$12.50
MistralMistral Medium 3.5Open SourceApr 2026$3.30
DeepSeekDeepSeek V4 ProOpen SourceApr 2026$2.11
xAIGrok 4.3Apr 2026$1.63
NVIDIANemotron 3 Nano OmniCheapestOpen SourceApr 2026$0.00
AnthropicClaude Opus 4.6Mar 2026$11.00
AnthropicClaude Sonnet 4.6Mar 2026$6.60
MistralMistral SmallOpen SourceMar 2026$0.285
MistralMistral LargeOpen SourceDec 2025$0.800
AnthropicClaude Haiku 4.5Oct 2025$2.20
MetaLlama 4 ScoutOpen SourceApr 2025$0.00
MetaLlama 4 MaverickOpen SourceApr 2025$0.00
GoogleGemini 2.5 ProMar 2025$3.88
OpenAIo3-miniJan 2025$2.09
OpenAIo1Dec 2024$28.50
OpenAIGPT-4o-miniJul 2024$0.285
OpenAIGPT-4oMay 2024$4.75
CohereCommand R+Apr 2024$4.75
CohereCommand RMar 2024$0.800

Cost breakdown for Nemotron 3.5 Lightning:

700K input tokens x $0.08/1M + 300K output tokens x $0.20/1M = $0.116/month

Frequently Asked Questions

How much does the Claude API cost?▾
Claude API pricing varies by model tier. Claude Fable 5.1, the top generally available tier, costs $10/1M input tokens and $50/1M output tokens. Claude Opus 5.5, released September 22, 2026 and now Anthropic's recommended starting point for most workloads, is $4/1M input and $20/1M output with cache reads at $0.20. Claude Opus 5 remains available at $5/1M input and $25/1M output, and Claude Sonnet 5 is $2/1M input and $10/1M output; all four ship with a 1M context window. Claude Haiku 4.5 is the most affordable at $1/1M input and $5/1M output with a 200K context window.
What is the cheapest AI API?▾
In our pricing data, the cheapest paid APIs include NVIDIA Nemotron 3.5 Lightning at $0.08/1M input tokens, and poolside Laguna S 2.1 and OpenAI GPT-6 Luna at $0.10/1M input. Google shut down Gemini 2.0 Flash, once a $0.10 budget default, on June 1, 2026. Open-weight models like DeepSeek V4.1 Flash and Qwen3.8-27B are free to self-host, though you will pay for compute infrastructure.
How are AI API tokens counted?▾
AI tokens are the basic units of text that language models process. One token is roughly 4 characters or about 0.75 words in English. A 1,000-word article is approximately 1,333 tokens. Pricing is typically quoted per 1 million tokens, with input (prompt) and output (completion) priced separately.
Which AI API is best for production?▾
The best AI API for production depends on your use case. Claude Sonnet 5 and GPT-6 Sol offer strong all-around performance at the same $2/$10 rate. For budget-sensitive applications, GPT-6 Luna and Gemini 3.8 Flash deliver solid results at a fraction of the price. For tasks requiring deep reasoning, Claude Opus 5.5 leads the independent Vals and Artificial Analysis indexes at $4/$20, and Claude Fable 5.1 or GPT-6 Astra are top choices at $10/$50.