What every major AI model actually costs, per million tokens.
Pulled from each provider’s official pricing page and reviewed by hand before publishing — not an unattended scraper. Last verified 2026-07-27.
Workload cost calculator
Estimate monthly cost across every tracked model for a workload you describe -- start from a preset, then adjust the numbers to match your actual usage.
Short conversational replies with a system prompt and recent chat history.
Batch rates aren’t modeled yet for these providers — real-time pricing is used either way.
| Model | Provider | Monthly cost | Fit | Input | Output | Context | Speed | Status |
|---|---|---|---|---|---|---|---|---|
Qwen3.7 Flashtiered | Qwen (Alibaba Model Studio) | $2.45 | Compatible | $0.03 | $0.13 | 1M | fast | production |
DeepSeek V4 Flash | DeepSeek | $6.96 | Compatible | $0.14 | $0.28 | 1M | fast | production |
GPT-5.4 Nano | OpenAI | $21.15 | Compatible | $0.2 | $1.25 | 400K | fast | production |
Qwen3.7 Plustiered | Qwen (Alibaba Model Studio) | $21.38 | Compatible | $0.276 | $1.10 | 1M | standard | production |
DeepSeek V4 Pro | DeepSeek | $21.45 | Compatible | $0.435 | $0.87 | 1M | standard | production |
Qwen3.6 Flashtiered | Qwen (Alibaba Model Studio) | $25.65 | Compatible | $0.25 | $1.50 | 1M | fast | production |
Moonshot V1 8K | Kimi (Moonshot) | $37.00 | Compatible | $0.2 | $2.00 | 8K | standard | production |
Moonshot V1 8K Visionpreview | Kimi (Moonshot) | $37.00 | Compatible | $0.2 | $2.00 | 8K | standard | preview |
Gemini 2.5 Flash | $39.53 | Compatible | $0.3 | $2.50 | 1M | fast | production | |
Kimi K2.5 | Kimi (Moonshot) | $55.50 | Compatible | $0.6 | $3.00 | 262K | standard | production |
Grok Build 0.1tieredpreview | xAI | $56.20 | Compatible | $1.00 | $2.00 | 256K | fast | preview |
Grok 4.20 0309 Non-Reasoningtiered | xAI | $68.45 | Compatible | $1.25 | $2.50 | 1M | fast | production |
Grok 4.20 0309 Reasoningtiered | xAI | $68.45 | Compatible | $1.25 | $2.50 | 1M | standard | production |
Grok 4.20 Multi-Agent 0309tieredpreview | xAI | $68.45 | Compatible | $1.25 | $2.50 | 1M | slow | preview |
Grok 4.3tiered | xAI | $68.45 | Compatible | $1.25 | $2.50 | 1M | standard | production |
GPT-5.4 Mini | OpenAI | $76.95 | Compatible | $0.75 | $4.50 | 400K | fast | production |
Kimi K2.6 | Kimi (Moonshot) | $78.56 | Compatible | $0.95 | $4.00 | 262K | standard | production |
Kimi K2.7 Code | Kimi (Moonshot) | $79.64 | Compatible | $0.95 | $4.00 | 262K | standard | production |
Moonshot V1 32K | Kimi (Moonshot) | $97.50 | Compatible | $1.00 | $3.00 | 33K | standard | production |
Moonshot V1 32K Visionpreview | Kimi (Moonshot) | $97.50 | Compatible | $1.00 | $3.00 | 33K | standard | preview |
GPT-5.6 Lunatiered | OpenAI | $102.60 | Compatible | $1.00 | $6.00 | 1.1M | fast | production |
Qwen3.7 Max | Qwen (Alibaba Model Studio) | $107.43 | Compatible | $1.65 | $4.95 | 1M | standard | production |
Grok 4.5tiered | xAI | $133.80 | Compatible | $2.00 | $6.00 | 500K | standard | production |
Kimi K2.7 Code HighSpeed | Kimi (Moonshot) | $159.28 | Compatible | $1.90 | $8.00 | 262K | fast | production |
Moonshot V1 128K | Kimi (Moonshot) | $182.50 | Compatible | $2.00 | $5.00 | 131K | standard | production |
Moonshot V1 128K Visionpreview | Kimi (Moonshot) | $182.50 | Compatible | $2.00 | $5.00 | 131K | standard | preview |
GPT-5.3 Codex | OpenAI | $223.30 | Compatible | $1.75 | $14.00 | 400K | standard | production |
GPT-5.4tiered | OpenAI | $256.50 | Compatible | $2.50 | $15.00 | 1.1M | standard | production |
GPT-5.6 Terratiered | OpenAI | $256.50 | Compatible | $2.50 | $15.00 | 1.1M | fast | production |
Kimi K3 | Kimi (Moonshot) | $270.30 | Compatible | $3.00 | $15.00 | 1.0M | slow | production |
Chat Latest | OpenAI | $513.00 | Compatible | $5.00 | $30.00 | 400K | fast | production |
GPT-5.5tiered | OpenAI | $513.00 | Compatible | $5.00 | $30.00 | 1.1M | standard | production |
GPT-5.6 Soltiered | OpenAI | $513.00 | Compatible | $5.00 | $30.00 | 1.1M | standard | production |
GPT-5.4 Protiered | OpenAI | $4,050.00 | Compatible | $30.00 | $180.00 | 1.1M | slow | production |
GPT-5.5 Protiered | OpenAI | $4,050.00 | Compatible | $30.00 | $180.00 | 1.1M | slow | production |
Ministral 3 — 3B | Mistral | $7.25 | Partial / unknown | $0.1 | $0.1 | — | fast | production |
Gemini 2.5 Flash-Lite | $7.76 | Partial / unknown | $0.1 | $0.4 | — | fast | production | |
Devstral Small 2preview | Mistral | $9.75 | Partial / unknown | $0.1 | $0.3 | — | fast | preview |
Ministral 3 — 8B | Mistral | $10.88 | Partial / unknown | $0.15 | $0.15 | — | fast | production |
Mistral NeMo | Mistral | $10.88 | Partial / unknown | $0.15 | $0.15 | — | fast | production |
Ministral 3 — 14B | Mistral | $14.50 | Partial / unknown | $0.2 | $0.2 | — | fast | production |
Mistral Small 4 | Mistral | $16.50 | Partial / unknown | $0.15 | $0.6 | — | fast | production |
Codestral | Mistral | $29.25 | Partial / unknown | $0.3 | $0.9 | — | fast | production |
Gemini 3.5 Flash-Lite | $39.53 | Partial / unknown | $0.3 | $2.50 | — | fast | production | |
Magistral Small | Mistral | $48.75 | Partial / unknown | $0.5 | $1.50 | — | standard | production |
Mistral Large 3 | Mistral | $48.75 | Partial / unknown | $0.5 | $1.50 | — | standard | production |
Devstral 2 | Mistral | $49.00 | Partial / unknown | $0.4 | $2.00 | — | standard | production |
Mixtral 8x7B | Mistral | $50.75 | Partial / unknown | $0.7 | $0.7 | — | standard | production |
Claude Haiku 4.5 | Anthropic | $90.10 | Partial / unknown | $1.00 | $5.00 | — | fast | production |
Gemini 3.6 Flash | $135.15 | Partial / unknown | $1.50 | $7.50 | — | fast | production | |
Gemini 3.5 Flash | $153.90 | Partial / unknown | $1.50 | $9.00 | — | fast | production | |
Gemini 2.5 Protiered | $159.50 | Partial / unknown | $1.25 | $10.00 | — | standard | production | |
Claude Sonnet 5 (introductory) | Anthropic | $180.20 | Partial / unknown | $2.00 | $10.00 | — | fast | production |
Magistral Medium | Mistral | $182.50 | Partial / unknown | $2.00 | $5.00 | — | slow | production |
Mistral Medium 3.5 | Mistral | $183.75 | Partial / unknown | $1.50 | $7.50 | — | standard | production |
Mixtral 8x22B | Mistral | $195.00 | Partial / unknown | $2.00 | $6.00 | — | standard | production |
Gemini 3.1 Pro Previewtieredpreview | $205.20 | Partial / unknown | $2.00 | $12.00 | — | standard | preview | |
Claude Sonnet 5 (standard) | Anthropic | $270.30 | Partial / unknown | $3.00 | $15.00 | — | fast | production |
Claude Opus 5 | Anthropic | $450.50 | Partial / unknown | $5.00 | $25.00 | — | standard | production |
Claude Fable 5 | Anthropic | $901.00 | Partial / unknown | $10.00 | $50.00 | — | slow | production |
Claude Opus 4.1 | Anthropic | $1,351.50 | Partial / unknown | $15.00 | $75.00 | — | slow | deprecated |
Claude Haiku 3.5 | Anthropic | $72.08 | Incompatible | $0.8 | $4.00 | — | fast | retired |
Assumptions
- Standard published prices are used; negotiated or enterprise pricing is not reflected.
- Provider-specific tiers (request-length, promotional, regional) are not yet calculated -- tiered models are flagged, but the estimate uses their base rate.
- Batch discounts are only applied where a structured batch rate exists in the data -- currently none, so batch mode uses real-time rates.
- Tool-invocation and media-processing charges are excluded from this estimate.
- Real-world token usage and retries vary; treat this as a planning estimate, not a bill.
- Estimates exclude taxes and regional pricing uplifts.
Anthropic
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Claude Fable 5 Highest-capability long-running agents and demanding knowledge work textimagetool callingreasoningslow | production | $10.00 | $1.00 | $50.00 | — | link |
Claude Haiku 3.5 Fast, low-cost text and vision workloads textimagetool callingfast Retired except on Bedrock/Google Cloud | retired | $0.8 | $0.08 | $4.00 | — | link |
Claude Haiku 4.5 High-volume, latency-sensitive agents and assistants textimagetool callingreasoningfast | production | $1.00 | $0.1 | $5.00 | — | link |
Claude Opus 4.1 Complex analysis, coding, and long-running agent tasks textimagetool callingreasoningslow | deprecated | $15.00 | $1.50 | $75.00 | — | link |
Claude Opus 5 Complex agentic coding and enterprise knowledge work textimagetool callingreasoningstandard | production | $5.00 | $0.5 | $25.00 | — | link |
Claude Sonnet 5 (introductory) Balanced production coding, agents, and enterprise workflows textimagetool callingreasoningfast Introductory pricing, ends 2026-08-31 | production | $2.00 | $0.2 | $10.00 | — | link |
Claude Sonnet 5 (standard) Balanced production coding, agents, and enterprise workflows textimagetool callingreasoningfast | production | $3.00 | $0.3 | $15.00 | — | link |
DeepSeek
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
DeepSeek V4 Flash Fast, economical reasoning and straightforward agent tasks texttool callingreasoningfast | production | $0.14 | $0.0028 | $0.28 | 1M | link |
DeepSeek V4 Pro Advanced reasoning, STEM, coding, and agentic workflows texttool callingreasoningstandard | production | $0.435 | $0.0036 | $0.87 | 1M | link |
| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Gemini 2.5 Flash Fast multimodal processing and high-volume agent workflows textimageaudiovideotool callingreasoningfast Text/image/video input; audio input billed at $1.00 | production | $0.3 | $0.03 | $2.50 | 1M | link |
Gemini 2.5 Flash-Lite Low-cost, high-throughput classification and extraction textimageaudiovideotool callingreasoningfast Text/image/video input; audio input billed at $0.30 | production | $0.1 | $0.01 | $0.4 | — | link |
Gemini 2.5 Pro Complex reasoning, coding, and long multimodal documents textimageaudiovideotool callingreasoningstandard Price rises with prompt length; shown value is the lower tier -- see source for the upper tier ($2.50 / $0.25 / $15.00) | production | $1.25 | $0.125 | $10.00 | — | link |
Gemini 3.1 Pro Preview Frontier multimodal reasoning and complex agentic tasks textimageaudiovideotool callingreasoningstandard Price rises with prompt length; shown value is the lower tier -- see source for the upper tier ($4.00 / $0.40 / $18.00) | preview | $2.00 | $0.2 | $12.00 | — | link |
Gemini 3.5 Flash Agentic coding and sustained multimodal workflows textimageaudiovideotool callingreasoningfast | production | $1.50 | $0.15 | $9.00 | — | link |
Gemini 3.5 Flash-Lite Lowest-cost high-throughput execution in the Gemini 3.5 family textimageaudiovideotool callingreasoningfast | production | $0.3 | $0.03 | $2.50 | — | link |
Gemini 3.6 Flash Balanced agentic and multimodal tasks with efficient token use textimageaudiovideotool callingreasoningfast | production | $1.50 | $0.15 | $7.50 | — | link |
Kimi (Moonshot)
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Kimi K2.5 Cost-efficient multimodal reasoning and agent tasks textimagevideotool callingreasoningstandard Kimi notes that web_search is currently being updated and does not recommend it in the near term. | production | $0.6 | $0.1 | $3.00 | 262K | link |
Kimi K2.6needs review Coding agents, visual understanding, and flexible thinking workloads textimagevideotool callingreasoningstandard Kimi notes that web_search is currently being updated and does not recommend it in the near term. | production | $0.95 | $0.16 | $4.00 | 262K | link |
Kimi K2.7 Code Long-context programming, deep reasoning, and coding agents textimagevideotool callingreasoningstandard | production | $0.95 | $0.19 | $4.00 | 262K | link |
Kimi K2.7 Code HighSpeed High-speed long-context programming and coding agents textimagevideotool callingreasoningfast Same model as Kimi K2.7 Code, served at approximately 180 tokens/s and up to 260 tokens/s for short contexts. | production | $1.90 | $0.38 | $8.00 | 262K | link |
Kimi K3needs review Long-horizon coding, deep reasoning, and end-to-end knowledge work textimagevideotool callingreasoningslow Kimi notes that web_search is currently being updated and does not recommend it in the near term. | production | $3.00 | $0.3 | $15.00 | 1.0M | link |
Moonshot V1 128K Long-document text generation and analysis textstandard | production | $2.00 | — | $5.00 | 131K | link |
Moonshot V1 128K Vision Image understanding with long-document context textimagestandard | preview | $2.00 | — | $5.00 | 131K | link |
Moonshot V1 32K Medium-length documents and general text generation textstandard | production | $1.00 | — | $3.00 | 33K | link |
Moonshot V1 32K Vision Image understanding with medium-length context textimagestandard | preview | $1.00 | — | $3.00 | 33K | link |
Moonshot V1 8K Short-context general text generation textstandard | production | $0.2 | — | $2.00 | 8K | link |
Moonshot V1 8K Vision Short-context image understanding and text generation textimagestandard | preview | $0.2 | — | $2.00 | 8K | link |
Mistral
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Codestral Low-latency code completion, fill-in-the-middle, and generation textfast Premier low-latency coding model. | production | $0.3 | — | $0.9 | — | link |
Devstral 2 Autonomous software engineering and agentic coding texttool callingstandard Open-weights agentic coding model. | production | $0.4 | — | $2.00 | — | link |
Devstral Small 2 Lightweight multimodal coding agents textimagetool callingfast Labs model for lightweight coding agents. | preview | $0.1 | — | $0.3 | — | link |
Magistral Medium Domain-specific, transparent, and multilingual reasoning textimagetool callingreasoningslow Premier thinking model. | production | $2.00 | — | $5.00 | — | link |
Magistral Small Cost-efficient multilingual reasoning and multimodal analysis textimagetool callingreasoningstandard Premier lightweight thinking model. | production | $0.5 | — | $1.50 | — | link |
Ministral 3 — 14B Higher-quality edge inference and compact agent workloads texttool callingfast Open lightweight edge model. | production | $0.2 | — | $0.2 | — | link |
Ministral 3 — 3B Small edge deployments and lightweight agents texttool callingfast Open lightweight edge model. | production | $0.1 | — | $0.1 | — | link |
Ministral 3 — 8B Balanced edge inference and lightweight agents texttool callingfast Open lightweight edge model. | production | $0.15 | — | $0.15 | — | link |
Mistral Large 3 General-purpose multilingual and multimodal enterprise workloads textimagetool callingstandard Open-weight flagship multimodal and multilingual model. | production | $0.5 | — | $1.50 | — | link |
Mistral Medium 3.5 Enterprise reasoning, coding, agents, and multimodal work textimagetool callingreasoningstandard Open model. State-of-the-art performance with simplified enterprise deployment. | production | $1.50 | — | $7.50 | — | link |
Mistral NeMo Low-cost lightweight coding workloads textfast Open lightweight model trained specifically for code tasks. | production | $0.15 | — | $0.15 | — | link |
Mistral Small 4 Cost-efficient multilingual, multimodal, and agentic workloads textimagetool callingfast Open Apache 2.0 model. | production | $0.15 | — | $0.6 | — | link |
Mixtral 8x22B Higher-quality open-weight general text generation textstandard Open sparse Mixture-of-Experts model. | production | $2.00 | — | $6.00 | — | link |
Mixtral 8x7B Lightweight general text generation with open weights textstandard Open sparse Mixture-of-Experts model. | production | $0.7 | — | $0.7 | — | link |
OpenAI
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Chat Latest Testing the latest Instant model experience used in ChatGPT textimagetool callingfast Latest Instant model used in ChatGPT; its underlying snapshot is regularly updated. OpenAI recommends GPT-5.6 for production API usage. | production | $5.00 | $0.5 | $30.00 | 400K | link |
GPT-5.3 Codex Long-running software engineering and agentic coding textimagetool callingreasoningstandard Specialized Codex model for agentic software engineering. | production | $1.75 | $0.175 | $14.00 | 400K | link |
GPT-5.4 General complex reasoning, coding, and professional workflows textimagetool callingreasoningstandard Affordable model for coding and professional work. Prompts over 272K input tokens use long-context pricing for the full request: $5 input, $0.5 cached input, $22.5 output per 1M. Eligible regional-processing endpoints add 10%. | production | $2.50 | $0.25 | $15.00 | 1.1M | link |
GPT-5.4 Mini Cost-efficient production agents and routine reasoning tasks textimagetool callingreasoningfast Strong mini model for coding, computer use, and subagents. Eligible regional-processing endpoints add 10%. | production | $0.75 | $0.075 | $4.50 | 400K | link |
GPT-5.4 Nano High-volume classification, extraction, and simple automation textimagetool callingreasoningfast Lowest-cost GPT-5.4-class model for simple high-volume tasks. Eligible regional-processing endpoints add 10%. | production | $0.2 | $0.02 | $1.25 | 400K | link |
GPT-5.4 Pro Highest-quality GPT-5.4 reasoning for difficult professional work textimagetool callingreasoningslow Higher-compute GPT-5.4 model for maximum response quality. Prompts over 272K input tokens use long-context pricing for the full request: $60 input, $270 output per 1M. Eligible regional-processing endpoints add 10%. | production | $30.00 | — | $180.00 | 1.1M | link |
GPT-5.5 Complex production reasoning, coding, and tool-driven workflows textimagetool callingreasoningstandard Frontier coding and professional-work model. Prompts over 272K input tokens use long-context pricing for the full request: $10 input, $1 cached input, $45 output per 1M. Eligible regional-processing endpoints add 10%. | production | $5.00 | $0.5 | $30.00 | 1.1M | link |
GPT-5.5 Pro Maximum-quality GPT-5.5 reasoning and complex professional tasks textimagetool callingreasoningslow Higher-compute GPT-5.5 model for maximum response quality. Prompts over 272K input tokens use long-context pricing for the full request: $60 input, $270 output per 1M. Eligible regional-processing endpoints add 10%. | production | $30.00 | — | $180.00 | 1.1M | link |
GPT-5.6 Luna Cost-sensitive, high-volume production workloads textimagetool callingreasoningfast GPT-5.6 model for cost-sensitive, high-volume workloads. Prompts over 272K input tokens use long-context pricing for the full request: $2 input, $0.2 cached input, $2.5 cache write, $9 output per 1M. Eligible regional-processing endpoints add 10%. | production | $1.00 | $0.1 | $6.00 | 1.1M | link |
GPT-5.6 Sol Frontier complex reasoning, coding, and professional work textimagetool callingreasoningstandard Frontier GPT-5.6 model for complex professional work. Prompts over 272K input tokens use long-context pricing for the full request: $10 input, $1 cached input, $12.5 cache write, $45 output per 1M. Eligible regional-processing endpoints add 10%. | production | $5.00 | $0.5 | $30.00 | 1.1M | link |
GPT-5.6 Terra Strong intelligence with balanced cost and latency textimagetool callingreasoningfast GPT-5.6 model balancing intelligence and cost. Prompts over 272K input tokens use long-context pricing for the full request: $5 input, $0.5 cached input, $6.25 cache write, $22.5 output per 1M. Eligible regional-processing endpoints add 10%. | production | $2.50 | $0.25 | $15.00 | 1.1M | link |
Qwen (Alibaba Model Studio)
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Qwen3.6 Flash Low-cost multimodal understanding and fast agent workflows textimagevideotool callingreasoningfast International list price shown for requests up to 256K tokens. Above 256K through 1M: $1 input and $4 output. Explicit cache hits cost 10%; implicit hits cost 20%. | production | $0.25 | $0.025 | $1.50 | 1M | link |
Qwen3.7 Flash Ultra-low-cost, high-throughput reasoning and text generation texttool callingreasoningfast International list price shown for requests up to 32K tokens. 32K-256K: $0.10 input/$0.40 output; 256K-1M: $0.20 input/$0.80 output. Explicit cache hits cost 10%; implicit hits cost 20%. | production | $0.03 | $0.003 | $0.13 | 1M | link |
Qwen3.7 Max Highest-capability Qwen reasoning and text generation texttool callingreasoningstandard Global list price through 1M tokens. Explicit cache hits cost 10% and implicit cache hits 20% of input list price. Time-limited night/day discounts may apply. | production | $1.65 | $0.165 | $4.95 | 1M | link |
Qwen3.7 Plus Balanced multimodal reasoning, long video, and tool use textimagevideotool callingreasoningstandard Global list price shown for requests up to 256K tokens. Above 256K through 1M: $0.826 input and $3.301 output. Explicit cache hits cost 10%; implicit hits cost 20%. Time-limited night/day discounts may apply. | production | $0.276 | $0.0276 | $1.10 | 1M | link |
xAI
USD · per 1M tokens| Model | Status | Input | Cached input | Output | Context | Source |
|---|---|---|---|---|---|---|
Grok 4.20 0309 Non-Reasoning Fast general-purpose responses and tool-driven workflows textimagetool callingfast Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $2.5 input, $0.4 cached input, and $5 output per 1M. Batch API discount: 20%. Priority processing costs 2x standard rates. | production | $1.25 | $0.2 | $2.50 | 1M | link |
Grok 4.20 0309 Reasoning Deep reasoning, analysis, and tool-driven workflows textimagetool callingreasoningstandard Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $2.5 input, $0.4 cached input, and $5 output per 1M. Batch API discount: 20%. Priority processing costs 2x standard rates. | production | $1.25 | $0.2 | $2.50 | 1M | link |
Grok 4.20 Multi-Agent 0309 Parallel multi-agent deep research and complex investigations textimagetool callingreasoningslow Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $2.5 input, $0.4 cached input, and $5 output per 1M. Batch API discount: 20%. Priority processing costs 2x standard rates. | preview | $1.25 | $0.2 | $2.50 | 1M | link |
Grok 4.3 General reasoning and tool-driven production workflows textimagetool callingreasoningstandard Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $2.5 input, $0.4 cached input, and $5 output per 1M. Batch API discount: 20%. Priority processing costs 2x standard rates. | production | $1.25 | $0.2 | $2.50 | 1M | link |
Grok 4.5 Agentic software engineering and complex coding workflows textimagetool callingreasoningstandard Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $4 input, $0.6 cached input, and $12 output per 1M. Priority processing costs 2x standard rates. | production | $2.00 | $0.3 | $6.00 | 500K | link |
Grok Build 0.1 Fast agentic coding, debugging, web development, and MCP texttool callingfast Short-context prices shown. At prompts of 200K tokens or more, all tokens cost $2 input, $0.4 cached input, and $4 output per 1M. Priority processing costs 2x standard rates. | preview | $1.00 | $0.2 | $2.00 | 256K | link |
Sources: OpenAI, Anthropic, Google, DeepSeek, Qwen (Alibaba Model Studio), Kimi (Moonshot), Mistral, xAI. Prices are informational and may lag behind provider changes between refreshes — always confirm against the linked source before billing decisions.
