AI API Token Pricing Comparison
Compare pay-as-you-go API prices in USD per 1 million tokens, updated ; for monthly ChatGPT, Claude, Gemini, and Copilot plans, use subscription pricing, and verify rates directly before purchase because providers can change them without notice.
| Try it | |||||||
|---|---|---|---|---|---|---|---|
GPT-5.6 Sol | OpenAI | High | 1.1M | $4.00 | $0.40 | $20.00 | Try API → |
GPT-5.6 Terra | OpenAI | Mid | 1.1M | $2.00 | $0.20 | $12.00 | Try API → |
GPT-5.6 Luna | OpenAI | Low | 1.1M | $0.20 | $0.02 | $1.20 | Try API → |
Claude Fable 5 | Anthropic | High | 1M | $10.00 | $1.00 | $50.00 | Try API → |
Claude Opus 5 | Anthropic | High | 1M | $5.00 | $0.50 | $25.00 | Try API → |
Claude Sonnet 5 | Anthropic | Mid | 1M | $2.00 | $0.20 | $10.00 | Try API → |
Gemini 3.1 ProPreview | High | - | $2.00 | $0.20 | $12.00 | Try in AI Studio → | |
DeepSeek V4 Pro 0813DeepSeek | 3 providers | High | - | $0.66 | $0.022 | $1.98 | Try on Novita → |
GLM-5.2Z.ai | 3 providers | Mid | 1M | $1.40 | $0.26 | $4.40 | Try on Novita → |
Grok 4.6 | xAI | High | 500,000 | $2.00 | $0.50 | $6.00 | Try API → |
Llama 4 MaverickMeta | 2 providers | High | 1M | $0.15 | - | $0.60 | Try on Novita → |
GPT-OSS Safeguard 20BPreviewOpenAI | Groq | High | - | $0.075 | - | $0.30 | Run on Vast → |
GPT-5.5 | OpenAI | High | - | $5.00 | $0.50 | $30.00 | Try API → |
GPT-5.5 Cyber | OpenAI | High | - | $12.50 | $1.25 | $75.00 | Try API → |
GPT-5.6 Cyber | OpenAI | High | - | $12.50 | $1.25 | $75.00 | Try API → |
GPT-5.4 Pro | OpenAI | High | - | $30.00 | - | $180.00 | Try API → |
GPT-5.5 Pro | OpenAI | High | - | $30.00 | - | $180.00 | Try API → |
GPT-5.4 | OpenAI | Mid | - | $2.50 | $0.25 | $15.00 | Try API → |
GPT-OSS 20BOpenAI | 2 providers | Low | - | $0.05 | $0.0375 | $0.20 | Run on Vast → |
GPT-OSS 120BOpenAI | 2 providers | Low | - | $0.15 | $0.075 | $0.60 | Run on Vast → |
GPT-5.4 nano | OpenAI | Low | - | $0.20 | $0.02 | $1.25 | Try API → |
GPT-5.4 mini | OpenAI | Low | - | $0.75 | $0.075 | $4.50 | Try API → |
Claude Mythos 5Preview | Anthropic | High | 1M | $10.00 | $1.00 | $50.00 | Try API → |
Claude Haiku 4.5 | Anthropic | Low | - | $1.00 | $0.10 | $5.00 | Try API → |
Gemma 4 31B IT PearlGoogle | Together | Mid | - | $0.28 | - | $0.86 | Run on Vast → |
Gemma 4 31B ITGoogle | Together | Mid | - | $0.39 | - | $0.97 | Run on Vast → |
Gemma 3n E4B InstructGoogle | Together | Low | - | $0.06 | - | $0.12 | Run on Vast → |
Gemini 3.1 Flash-LitePreview | Low | - | $0.25 | $0.025 | $1.50 | Try in AI Studio → | |
Gemini 3 FlashPreview | Low | - | $0.50 | $0.05 | $3.00 | Try in AI Studio → | |
Gemini 3.7 Flash | Low | - | $0.75 | $0.075 | $3.75 | Try in AI Studio → | |
DeepSeek V3.1 TerminusDeepSeek | Novita | High | - | $0.27 | $0.135 | $1.00 | Try on Novita → |
DeepSeek R1 (Turbo)DeepSeek | Novita | High | - | $0.70 | $0.35 | $2.50 | Try on Novita → |
DeepSeek R1 Distill Llama 70BDeepSeek | Novita | High | - | $0.80 | - | $0.80 | Try on Novita → |
DeepSeek V4 Pro 0813DeepSeek | Novita | High | - | $1.32 | $0.132 | $3.96 | Try on Novita → |
DeepSeek-OCR 2DeepSeek | Novita | Mid | - | $0.03 | - | $0.03 | Try on Novita → |
DeepSeek V3.2DeepSeek | Novita | Mid | - | $0.269 | $0.1345 | $0.40 | Try on Novita → |
DeepSeek V3 0324DeepSeek | Novita | Mid | - | $0.27 | $0.135 | $1.12 | Try on Novita → |
DeepSeek V3.1DeepSeek | Novita | Mid | - | $0.27 | $0.135 | $1.00 | Try on Novita → |
DeepSeek V3.2 ExpDeepSeek | Novita | Mid | - | $0.27 | - | $0.41 | Try on Novita → |
DeepSeek V4 Flash 0731DeepSeek | 2 providers | Low | - | $0.14 | $0.007 | $0.28 | Try on Novita → |
DeepSeek V4 Flash 0731DeepSeek | Novita | Low | - | $0.44 | $0.028 | $1.32 | Try on Novita → |
GLM-5Z.ai | Novita | High | - | $1.00 | $0.20 | $3.20 | Try on Novita → |
GLM-5.1Z.ai | 2 providers | High | - | $1.38 | $0.26 | $4.40 | Try on Novita → |
AutoGLM-Phone-9B-MultilingualZ.ai | Novita | Mid | - | $0.035 | - | $0.138 | Try on Novita → |
GLM 5.3Z.ai | Novita | Mid | - | $1.40 | $0.26 | $4.40 | Try on Novita → |
GLM-5.3 | Z.ai | Mid | 1M | $1.40 | $0.26 | $4.40 | Coding plan → |
Grok Build 0.1 | xAI | Mid | 256,000 | $1.00 | $0.20 | $2.00 | Try API → |
Grok 4.3 | xAI | Mid | 1M | $1.25 | $0.20 | $2.50 | Try API → |
Grok 4.20 0309 Non-Reasoning | xAI | Mid | 1M | $1.25 | $0.20 | $2.50 | Try API → |
Mistral NeMoMistral | Novita | Mid | - | $0.04 | - | $0.17 | Try on Novita → |
Ministral 3B | Mistral | Low | - | $0.10 | - | $0.10 | Try API → |
Ministral 8B | Mistral | Low | - | $0.15 | - | $0.15 | Try API → |
Mistral Small 4 | Mistral | Low | - | $0.15 | - | $0.60 | Try API → |
Llama 3.1 8B InstructMeta | 2 providers | Mid | - | $0.02 | - | $0.05 | Try on Novita → |
Llama 4 ScoutMeta | 2 providers | Mid | 10M | $0.10 | - | $0.30 | Try on Novita → |
Llama 3.3 70B InstructMeta | 3 providers | Mid | - | $0.135 | - | $0.40 | Try on Novita → |
Llama 3 8B Instruct LiteMeta | Together | Mid | - | $0.14 | - | $0.14 | Run on Vast → |
Nemotron 3 Ultra 550B A55BNVIDIA | Together | High | - | $0.60 | $0.20 | $3.60 | Run on Vast → |
Qwen2.5 7B Instruct TurboAlibaba | Together | Mid | - | $0.30 | - | $0.30 | Try API → |
DCCogito v2.1 671BDeep Cogito | Together | Mid | - | $1.25 | - | $1.25 | Run on Vast → |
EARNJ-1 InstructEssential AI | Together | Low | - | $0.15 | - | $0.15 | Run on Vast → |
Qwen3.5 9BAlibaba | Together | Low | - | $0.17 | - | $0.25 | Try API → |
Command A | Cohere | High | - | $2.50 | - | $10.00 | Try API → |
Embed v3 English | Cohere | Mid | - | $0.10 | - | $0.00 | Try API → |
Embed v3 Multilingual | Cohere | Mid | - | $0.10 | - | $0.00 | Try API → |
Rerank v3 | Cohere | Mid | - | $2.00 | - | $0.00 | Try API → |
Command R7B | Cohere | Low | - | $0.0375 | - | $0.15 | Try API → |
Sonar Deep Research | Perplexity | High | - | $2.00 | - | $8.00 | Try API → |
Sonar Reasoning Pro | Perplexity | High | - | $2.00 | - | $8.00 | Try API → |
Sonar Pro | Perplexity | High | - | $3.00 | - | $15.00 | Try API → |
Sonar | Perplexity | Low | - | $1.00 | - | $1.00 | Try API → |
Qwen3 235B A22B Instruct 2507Alibaba | 2 providers | Mid | - | $0.09 | - | $0.58 | Try on Novita → |
Qwen3 235B A22BAlibaba | Novita | High | - | $0.20 | - | $0.80 | Try on Novita → |
MiniMax M2MiniMax | Novita | High | - | $0.30 | $0.03 | $1.20 | Try on Novita → |
MiniMax M2.1MiniMax | Novita | High | - | $0.30 | $0.03 | $1.20 | Try on Novita → |
MiniMax M2.5MiniMax | Novita | High | - | $0.30 | $0.03 | $1.20 | Try on Novita → |
MiniMax M2.7MiniMax | 2 providers | Mid | - | $0.30 | $0.06 | $1.20 | Try on Novita → |
MiniMax M3MiniMax | 2 providers | High | 1M | $0.30 | $0.06 | $1.20 | Try API → |
Qwen3 235B A22B Thinking 2507Alibaba | Novita | High | - | $0.30 | - | $3.00 | Try on Novita → |
Qwen3 VL 235B A22B InstructAlibaba | Novita | High | - | $0.30 | - | $1.50 | Try on Novita → |
Qwen3 Coder 480B A35B InstructAlibaba | Novita | High | - | $0.38 | - | $1.55 | Try on Novita → |
MiniMax M1MiniMax | Novita | High | - | $0.55 | - | $2.20 | Try on Novita → |
Qwen3.5-397B-A17BAlibaba | 2 providers | Mid | - | $0.60 | $0.35 | $3.60 | Try on Novita → |
Qwen3.7-MaxAlibaba | 2 providers | High | 1M | $1.25 | $0.125 | $3.75 | Try on Novita → |
Kimi K3Moonshot | 3 providers | High | 1M | $2.70 | $0.27 | $13.50 | Try on Novita → |
Qwen3 Coder 30B A3B InstructAlibaba | Novita | Mid | - | $0.07 | - | $0.27 | Try on Novita → |
Qwen3 Next 80B A3B InstructAlibaba | Novita | Mid | - | $0.15 | - | $1.50 | Try on Novita → |
Qwen3 Coder NextAlibaba | Novita | Mid | - | $0.20 | - | $1.50 | Try on Novita → |
Qwen3 VL 30B A3B InstructAlibaba | Novita | Mid | - | $0.20 | - | $0.70 | Try on Novita → |
Qwen3.6-35B-A3BAlibaba | Novita | Mid | - | $0.248 | - | $1.49 | Try on Novita → |
Qwen MT PlusAlibaba | Novita | Mid | - | $0.25 | - | $0.75 | Try on Novita → |
Qwen3.5-35B-A3BAlibaba | Novita | Mid | - | $0.25 | - | $2.00 | Try on Novita → |
Qwen3.5-27BAlibaba | Novita | Mid | - | $0.30 | - | $2.40 | Try on Novita → |
Qwen3.7-Plus | Alibaba | Mid | 1M | $0.32 | $0.032 | $1.28 | Try API → |
Qwen 2.5 72B InstructAlibaba | Novita | Mid | - | $0.38 | - | $0.40 | Try on Novita → |
Qwen3.5-122B-A10BAlibaba | Novita | Mid | - | $0.40 | - | $3.20 | Try on Novita → |
Kimi K2 InstructMoonshot | Novita | Mid | - | $0.57 | $0.15 | $2.30 | Try on Novita → |
Kimi K2.5Moonshot | Novita | Mid | - | $0.60 | $0.10 | $3.00 | Try on Novita → |
Qwen3.6-27BAlibaba | Novita | Mid | - | $0.60 | - | $3.60 | Try on Novita → |
Kimi K2.6Moonshot | 2 providers | Mid | - | $0.80 | $0.16 | $3.40 | Try on Novita → |
Kimi K2.7 CodeMoonshot | 2 providers | Mid | - | $0.95 | $0.19 | $4.00 | Try on Novita → |
Qwen3.8 MaxAlibaba | Novita | Mid | - | $2.00 | $0.25 | $6.00 | Try on Novita → |
Qwen3.5 4BAlibaba | Ccastform | Low | - | $0.03 | - | $0.15 | Open provider → |
Qwen3.6-Flash | Alibaba | Low | 1M | $0.25 | $0.025 | $1.50 | Try API → |
- GPT-5.6 SolOpenAIHigh
- Input
- $4.00
- Cached
- $0.40
- Output
- $20.00
- GPT-5.6 TerraOpenAIMid
- Input
- $2.00
- Cached
- $0.20
- Output
- $12.00
- GPT-5.6 LunaOpenAILow
- Input
- $0.20
- Cached
- $0.02
- Output
- $1.20
- Claude Fable 5AnthropicHigh
- Input
- $10.00
- Cached
- $1.00
- Output
- $50.00
- Claude Opus 5AnthropicHigh
- Input
- $5.00
- Cached
- $0.50
- Output
- $25.00
- Claude Sonnet 5AnthropicMid
- Input
- $2.00
- Cached
- $0.20
- Output
- $10.00
- Gemini 3.1 ProPreviewGoogleHigh
- Input
- $2.00
- Cached
- $0.20
- Output
- $12.00
- DeepSeek V4 Pro 0813DeepSeek3 providersHigh
- Input
- $0.66
- Cached
- $0.022
- Output
- $1.98
- GLM-5.2Z.ai3 providersMid
- Input
- $1.40
- Cached
- $0.26
- Output
- $4.40
- Grok 4.6xAIHigh
- Input
- $2.00
- Cached
- $0.50
- Output
- $6.00
- Llama 4 MaverickMeta2 providersHigh
- Input
- $0.15
- Cached
- -
- Output
- $0.60
- GPT-OSS Safeguard 20BPreviewOpenAIGroqHigh
- Input
- $0.075
- Cached
- -
- Output
- $0.30
- GPT-5.5OpenAIHigh
- Input
- $5.00
- Cached
- $0.50
- Output
- $30.00
- GPT-5.5 CyberOpenAIHigh
- Input
- $12.50
- Cached
- $1.25
- Output
- $75.00
- GPT-5.6 CyberOpenAIHigh
- Input
- $12.50
- Cached
- $1.25
- Output
- $75.00
- GPT-5.4 ProOpenAIHigh
- Input
- $30.00
- Cached
- -
- Output
- $180.00
- GPT-5.5 ProOpenAIHigh
- Input
- $30.00
- Cached
- -
- Output
- $180.00
- GPT-5.4OpenAIMid
- Input
- $2.50
- Cached
- $0.25
- Output
- $15.00
- GPT-OSS 20BOpenAI2 providersLow
- Input
- $0.05
- Cached
- $0.0375
- Output
- $0.20
- GPT-OSS 120BOpenAI2 providersLow
- Input
- $0.15
- Cached
- $0.075
- Output
- $0.60
- GPT-5.4 nanoOpenAILow
- Input
- $0.20
- Cached
- $0.02
- Output
- $1.25
- GPT-5.4 miniOpenAILow
- Input
- $0.75
- Cached
- $0.075
- Output
- $4.50
- Claude Mythos 5PreviewAnthropicHigh
- Input
- $10.00
- Cached
- $1.00
- Output
- $50.00
- Claude Haiku 4.5AnthropicLow
- Input
- $1.00
- Cached
- $0.10
- Output
- $5.00
- Gemma 4 31B IT PearlGoogleTogetherMid
- Input
- $0.28
- Cached
- -
- Output
- $0.86
- Gemma 4 31B ITGoogleTogetherMid
- Input
- $0.39
- Cached
- -
- Output
- $0.97
- Gemma 3n E4B InstructGoogleTogetherLow
- Input
- $0.06
- Cached
- -
- Output
- $0.12
- Gemini 3.1 Flash-LitePreviewGoogleLow
- Input
- $0.25
- Cached
- $0.025
- Output
- $1.50
- Gemini 3 FlashPreviewGoogleLow
- Input
- $0.50
- Cached
- $0.05
- Output
- $3.00
- Gemini 3.7 FlashGoogleLow
- Input
- $0.75
- Cached
- $0.075
- Output
- $3.75
- DeepSeek V3.1 TerminusDeepSeekNovitaHigh
- Input
- $0.27
- Cached
- $0.135
- Output
- $1.00
- DeepSeek R1 (Turbo)DeepSeekNovitaHigh
- Input
- $0.70
- Cached
- $0.35
- Output
- $2.50
- DeepSeek R1 Distill Llama 70BDeepSeekNovitaHigh
- Input
- $0.80
- Cached
- -
- Output
- $0.80
- DeepSeek V4 Pro 0813DeepSeekNovitaHigh
- Input
- $1.32
- Cached
- $0.132
- Output
- $3.96
- DeepSeek-OCR 2DeepSeekNovitaMid
- Input
- $0.03
- Cached
- -
- Output
- $0.03
- DeepSeek V3.2DeepSeekNovitaMid
- Input
- $0.269
- Cached
- $0.1345
- Output
- $0.40
- DeepSeek V3 0324DeepSeekNovitaMid
- Input
- $0.27
- Cached
- $0.135
- Output
- $1.12
- DeepSeek V3.1DeepSeekNovitaMid
- Input
- $0.27
- Cached
- $0.135
- Output
- $1.00
- DeepSeek V3.2 ExpDeepSeekNovitaMid
- Input
- $0.27
- Cached
- -
- Output
- $0.41
- DeepSeek V4 Flash 0731DeepSeek2 providersLow
- Input
- $0.14
- Cached
- $0.007
- Output
- $0.28
- DeepSeek V4 Flash 0731DeepSeekNovitaLow
- Input
- $0.44
- Cached
- $0.028
- Output
- $1.32
- GLM-5Z.aiNovitaHigh
- Input
- $1.00
- Cached
- $0.20
- Output
- $3.20
- GLM-5.1Z.ai2 providersHigh
- Input
- $1.38
- Cached
- $0.26
- Output
- $4.40
- AutoGLM-Phone-9B-MultilingualZ.aiNovitaMid
- Input
- $0.035
- Cached
- -
- Output
- $0.138
- GLM 5.3Z.aiNovitaMid
- Input
- $1.40
- Cached
- $0.26
- Output
- $4.40
- GLM-5.3Z.aiMid
- Input
- $1.40
- Cached
- $0.26
- Output
- $4.40
- Grok Build 0.1xAIMid
- Input
- $1.00
- Cached
- $0.20
- Output
- $2.00
- Grok 4.3xAIMid
- Input
- $1.25
- Cached
- $0.20
- Output
- $2.50
- Grok 4.20 0309 Non-ReasoningxAIMid
- Input
- $1.25
- Cached
- $0.20
- Output
- $2.50
- Mistral NeMoMistralNovitaMid
- Input
- $0.04
- Cached
- -
- Output
- $0.17
- Ministral 3BMistralLow
- Input
- $0.10
- Cached
- -
- Output
- $0.10
- Ministral 8BMistralLow
- Input
- $0.15
- Cached
- -
- Output
- $0.15
- Mistral Small 4MistralLow
- Input
- $0.15
- Cached
- -
- Output
- $0.60
- Llama 3.1 8B InstructMeta2 providersMid
- Input
- $0.02
- Cached
- -
- Output
- $0.05
- Llama 4 ScoutMeta2 providersMid
- Input
- $0.10
- Cached
- -
- Output
- $0.30
- Llama 3.3 70B InstructMeta3 providersMid
- Input
- $0.135
- Cached
- -
- Output
- $0.40
- Llama 3 8B Instruct LiteMetaTogetherMid
- Input
- $0.14
- Cached
- -
- Output
- $0.14
- Nemotron 3 Ultra 550B A55BNVIDIATogetherHigh
- Input
- $0.60
- Cached
- $0.20
- Output
- $3.60
- Qwen2.5 7B Instruct TurboAlibabaTogetherMid
- Input
- $0.30
- Cached
- -
- Output
- $0.30
- DCCogito v2.1 671BDeep CogitoTogetherMid
- Input
- $1.25
- Cached
- -
- Output
- $1.25
- EARNJ-1 InstructEssential AITogetherLow
- Input
- $0.15
- Cached
- -
- Output
- $0.15
- Qwen3.5 9BAlibabaTogetherLow
- Input
- $0.17
- Cached
- -
- Output
- $0.25
- Command ACohereHigh
- Input
- $2.50
- Cached
- -
- Output
- $10.00
- Embed v3 EnglishCohereMid
- Input
- $0.10
- Cached
- -
- Output
- $0.00
- Embed v3 MultilingualCohereMid
- Input
- $0.10
- Cached
- -
- Output
- $0.00
- Rerank v3CohereMid
- Input
- $2.00
- Cached
- -
- Output
- $0.00
- Command R7BCohereLow
- Input
- $0.0375
- Cached
- -
- Output
- $0.15
- Sonar Deep ResearchPerplexityHigh
- Input
- $2.00
- Cached
- -
- Output
- $8.00
- Sonar Reasoning ProPerplexityHigh
- Input
- $2.00
- Cached
- -
- Output
- $8.00
- Sonar ProPerplexityHigh
- Input
- $3.00
- Cached
- -
- Output
- $15.00
- SonarPerplexityLow
- Input
- $1.00
- Cached
- -
- Output
- $1.00
- Qwen3 235B A22B Instruct 2507Alibaba2 providersMid
- Input
- $0.09
- Cached
- -
- Output
- $0.58
- Qwen3 235B A22BAlibabaNovitaHigh
- Input
- $0.20
- Cached
- -
- Output
- $0.80
- MiniMax M2MiniMaxNovitaHigh
- Input
- $0.30
- Cached
- $0.03
- Output
- $1.20
- MiniMax M2.1MiniMaxNovitaHigh
- Input
- $0.30
- Cached
- $0.03
- Output
- $1.20
- MiniMax M2.5MiniMaxNovitaHigh
- Input
- $0.30
- Cached
- $0.03
- Output
- $1.20
- MiniMax M2.7MiniMax2 providersMid
- Input
- $0.30
- Cached
- $0.06
- Output
- $1.20
- MiniMax M3MiniMax2 providersHigh
- Input
- $0.30
- Cached
- $0.06
- Output
- $1.20
- Qwen3 235B A22B Thinking 2507AlibabaNovitaHigh
- Input
- $0.30
- Cached
- -
- Output
- $3.00
- Qwen3 VL 235B A22B InstructAlibabaNovitaHigh
- Input
- $0.30
- Cached
- -
- Output
- $1.50
- Qwen3 Coder 480B A35B InstructAlibabaNovitaHigh
- Input
- $0.38
- Cached
- -
- Output
- $1.55
- MiniMax M1MiniMaxNovitaHigh
- Input
- $0.55
- Cached
- -
- Output
- $2.20
- Qwen3.5-397B-A17BAlibaba2 providersMid
- Input
- $0.60
- Cached
- $0.35
- Output
- $3.60
- Qwen3.7-MaxAlibaba2 providersHigh
- Input
- $1.25
- Cached
- $0.125
- Output
- $3.75
- Kimi K3Moonshot3 providersHigh
- Input
- $2.70
- Cached
- $0.27
- Output
- $13.50
- Qwen3 Coder 30B A3B InstructAlibabaNovitaMid
- Input
- $0.07
- Cached
- -
- Output
- $0.27
- Qwen3 Next 80B A3B InstructAlibabaNovitaMid
- Input
- $0.15
- Cached
- -
- Output
- $1.50
- Qwen3 Coder NextAlibabaNovitaMid
- Input
- $0.20
- Cached
- -
- Output
- $1.50
- Qwen3 VL 30B A3B InstructAlibabaNovitaMid
- Input
- $0.20
- Cached
- -
- Output
- $0.70
- Qwen3.6-35B-A3BAlibabaNovitaMid
- Input
- $0.248
- Cached
- -
- Output
- $1.49
- Qwen MT PlusAlibabaNovitaMid
- Input
- $0.25
- Cached
- -
- Output
- $0.75
- Qwen3.5-35B-A3BAlibabaNovitaMid
- Input
- $0.25
- Cached
- -
- Output
- $2.00
- Qwen3.5-27BAlibabaNovitaMid
- Input
- $0.30
- Cached
- -
- Output
- $2.40
- Qwen3.7-PlusAlibabaMid
- Input
- $0.32
- Cached
- $0.032
- Output
- $1.28
- Qwen 2.5 72B InstructAlibabaNovitaMid
- Input
- $0.38
- Cached
- -
- Output
- $0.40
- Qwen3.5-122B-A10BAlibabaNovitaMid
- Input
- $0.40
- Cached
- -
- Output
- $3.20
- Kimi K2 InstructMoonshotNovitaMid
- Input
- $0.57
- Cached
- $0.15
- Output
- $2.30
- Kimi K2.5MoonshotNovitaMid
- Input
- $0.60
- Cached
- $0.10
- Output
- $3.00
- Qwen3.6-27BAlibabaNovitaMid
- Input
- $0.60
- Cached
- -
- Output
- $3.60
- Kimi K2.6Moonshot2 providersMid
- Input
- $0.80
- Cached
- $0.16
- Output
- $3.40
- Kimi K2.7 CodeMoonshot2 providersMid
- Input
- $0.95
- Cached
- $0.19
- Output
- $4.00
- Qwen3.8 MaxAlibabaNovitaMid
- Input
- $2.00
- Cached
- $0.25
- Output
- $6.00
- Qwen3.5 4BAlibabaCcastformLow
- Input
- $0.03
- Cached
- -
- Output
- $0.15
- Qwen3.6-FlashAlibabaLow
- Input
- $0.25
- Cached
- $0.025
- Output
- $1.50
Sponsored links may earn us a commission at no extra cost to you. Affiliate status never changes model ordering.
All product names, logos, and brands are property of their respective owners and are used for identification purposes only.
AI model price scale
Every current canonical model at its cheapest active host. USD per 1M tokens.
All product names, logos, and brands are property of their respective owners and are used for identification purposes only.
Which AI API is cheapest right now? We track 215 models across 17 providers. The cheapest flagship is GPT-5.6 Luna at $0.20 per 1M input tokens; the absolute cheapest production model is Llama 3.1 8B Instruct at $0.02 per 1M input. The most expensive we track is GPT-5.4 Pro at $30.00 input / 180.00 output. Download the raw data as JSON.
Novita open-model API
One API for Llama, Qwen, DeepSeek and GLM — new accounts get $100 in sandbox credits for 90 days.
Affiliate disclosure: sponsored links may earn us a commission at no extra cost to you. Affiliate status never changes model ordering.
Qwen model family
Compare every current Qwen route
Separate Alibaba first-party aliases from managed Qwen hosts, with cache and long-context tiers.
Need exact math?
Use the token cost calculator
Enter your input/output token volume and estimate monthly spend before choosing a model.
Labs benchmark
Compare cost per correct answer
See which models turn token spend into correct reasoning, extraction, coding, and pricing answers.
Model pages
Open detailed model pricing
See per-model calculators, cached-token math, price history, and alternatives for high-demand models.
Subscription or API?
Compare monthly plans vs API usage
If you use AI through a chat app, calculate whether a subscription is cheaper than raw API tokens.
Frequently asked questions
Quick answers about API token pricing, freshness, and how to compare providers.
What does this AI API pricing comparison cover?
This page tracks pay-as-you-go API token pricing for 215 AI models across 17 providers — OpenAI, Anthropic, Google, DeepSeek, Mistral, xAI, Cohere, Groq, Together AI, Perplexity, and Meta Llama hosts. Each row shows input price, output price, cached-input price (where the provider offers prompt caching), and batch-discounted rates per million tokens. The data snapshot shown is from 2026-08-25.
How often is the pricing data updated?
Prices are verified and reconciled daily by an automated pipeline that pulls from each provider's official pricing page and cross-checks against at least two independent sources before publishing. New models and price changes typically appear here within hours of the provider's announcement. If a provider hasn't moved its prices, the snapshot date stays the same — the current dataset is from 2026-08-25.
Which AI API is the cheapest right now?
As of 2026-08-25, the absolute cheapest production model we track is Llama 3.1 8B Instruct at $0.02 per million input tokens. The cheapest flagship-class model — meaning a top-tier model from a major lab, not a small or distilled variant — is GPT-5.6 Luna at $0.20 per million input tokens. The most expensive model we track is GPT-5.4 Pro at $30.00 input / 180.00 output per 1M tokens.
How do I compare two specific AI providers like OpenAI and Anthropic?
Use the table search and filter controls above to narrow to a single provider, then sort by input or output price. For deeper provider-level pages with FAQs, methodology, and historical pricing, see /openai-pricing/, /anthropic-pricing/, /google-pricing/, /deepseek-pricing/, /mistral-ai-pricing/, /cohere-pricing/, or any other provider listed. For workload-level math, the token cost calculator at /calculators/token-cost/ runs your input/output volume against every model side-by-side.
Should I pay per token via API or subscribe to ChatGPT Plus or Claude Pro?
It depends on volume and access pattern. Monthly subscriptions like ChatGPT Plus, Claude Pro, and Gemini Advanced at roughly $20/month are usually cheaper than API access for chat-style usage under about 5 million input tokens per month. Production workloads, agents, and anything that needs API access almost always cost less pay-as-you-go on the API. The /subscription-vs-api/ tool computes the exact break-even point for your specific volume.
Can I download the AI API pricing data as JSON?
Yes. The full live dataset is published at /api/pricing.json and updated whenever this page is. The schema is { id, name, family, provider, pricing: { inputPerM, cachedInputPerM, outputPerM }, status } for all 215 tracked models. The AI Pricing Guru dataset is provided for informational use. You may use it for personal, editorial, research, educational, and internal business purposes with appropriate attribution to AI Pricing Guru. Commercial redistribution, resale, republishing at scale, inclusion in a competing public dataset/API, or use as the primary data source for a commercial pricing-comparison product requires prior written permission.