Llama 4 Maverick (400B) Meta AI 1000000
💰 Total Cost Calculation (from Plugin)
Output: $0.003000
Output: $0.003000
Unit: $0.000000
Fees: $0.000000
Detailed Cost Analysis (from Plugin)
For 25,000 input tokens and 5,000 output tokens:
- Input Cost: $0.003750
- Output Cost: $0.003000
- Total Cost: $0.006750 (rounded ~ $0.01)
- Cost per 1K tokens: $0.000225
- Tokens per dollar: 4,444,444 tokens
- Context Window: 1000000 tokens
Speed & Performance Analysis
With a processing speed of 400 tokens per second and 150ms time to first token:
- Processing Time: 1 minute, 20.43 seconds
- Latency: 150 milliseconds to first token
- Base Throughput: 400 tokens/second
- Effective Throughput: 374 tokens/second (temperature-adjusted)
Best Use Cases
Want this applied to YOUR actual stack?
This calculator shows the math for Llama 4 Maverick (400B). Your decision needs more — current infrastructure, compliance requirements, actual workload patterns, volume tiers — that change which model is right for you.
Get a $39 personalized AI Architecture Audit. PDF tailored to your stack, delivered in under 60 seconds. 7-day no-questions-asked refund.
Get my instant AI audit — $39 →Claude Opus 4.6 Anthropic 1000000
💰 Total Cost Calculation (from Plugin)
Output: $0.031250 (rounded ~ $0.03)
Output: $0.031250 (rounded ~ $0.03)
Unit: $0.000000
Fees: $0.000000
Advanced Cost Breakdown (from Plugin)
Detailed Cost Analysis (from Plugin)
For 25,000 input tokens and 5,000 output tokens:
- Input Cost: $0.031250 (rounded ~ $0.03)
- Output Cost: $0.031250 (rounded ~ $0.03)
- Total Cost: $0.054063 (rounded ~ $0.05)
- Cost per 1K tokens: $0.001802
- Tokens per dollar: 554,913 tokens
- Context Window: 1000000 tokens
Speed & Performance Analysis
With a processing speed of 280 tokens per second and 380ms time to first token:
- Processing Time: 1 minute, 54.82 seconds
- Latency: 380 milliseconds to first token
- Base Throughput: 280 tokens/second
- Effective Throughput: 262 tokens/second (temperature-adjusted)
Best Use Cases
Want this applied to YOUR actual stack?
This calculator shows the math for Claude Opus 4.6. Your decision needs more — current infrastructure, compliance requirements, actual workload patterns, volume tiers — that change which model is right for you.
Get a $39 personalized AI Architecture Audit. PDF tailored to your stack, delivered in under 60 seconds. 7-day no-questions-asked refund.
Get my instant AI audit — $39 →✨ Market Recommendations AI Model Registry
← Back to Llama 4 Maverick (400B)| Rank | AI Model & Provider | Total Cost | vs Llama 4 Maverick (400B) | vs Claude Opus 4.6 |
|---|---|---|---|---|
| 🏆 |
Mistral Small 3
Mistral AI
|
$0.000831 Best Value | ↓ 87.7% cheaper | ↓ 98.5% cheaper |
| 🥈 |
Gemini 3.1 Flash Lite
Google
|
$0.003016 | ↓ 55.3% cheaper | ↓ 94.4% cheaper |
| 🥉 |
Mistral Large 3
Mistral AI
|
$0.004156 | ↓ 38.4% cheaper | ↓ 92.3% cheaper |
| #4 |
Gemini 3.5 Flash-Lite
Google
|
$0.004494 | ↓ 33.4% cheaper | ↓ 91.7% cheaper |
| #5 |
Gemini 2.5 Flash
Google
|
$0.004494 | ↓ 33.4% cheaper | ↓ 91.7% cheaper |
| #6 |
Gemini 3.8 Flash
Google
|
$0.008109 (rounded ~ $0.01) | ↑ 20.1% more | ↓ 85% cheaper |
| #7 |
GPT-5.4 mini
OpenAI
|
$0.009047 | ↑ 34% more | ↓ 83.3% cheaper |
| #8 |
o4-mini Deep Research
OpenAI
|
$0.009563 | ↑ 41.7% more | ↓ 82.3% cheaper |
| #9 |
o4-mini
OpenAI
|
$0.010519 | ↑ 55.8% more | ↓ 80.5% cheaper |
| #10 |
Claude Haiku 4.5
Anthropic
|
$0.010813 | ↑ 60.2% more | ↓ 80% cheaper |
| #11 |
Gemini 3.1 Flash
Google
|
$0.012063 (rounded ~ $0.01) | ↑ 78.7% more | ↓ 77.7% cheaper |
| #12 |
GPT-5.6 Luna
OpenAI
|
$0.012063 (rounded ~ $0.01) | ↑ 78.7% more | ↓ 77.7% cheaper |
| #13 |
Gemini 3.6 Flash
Google
|
$0.016219 (rounded ~ $0.02) | ↑ 140.3% more | ↓ 70% cheaper |
| #14 |
Gemini 3.5 Flash
Google
|
$0.018094 (rounded ~ $0.02) | ↑ 168.1% more | ↓ 66.5% cheaper |
| #15 |
Claude Sonnet 5
Anthropic
|
$0.021625 (rounded ~ $0.02) | ↑ 220.4% more | ↓ 60% cheaper |
| #16 |
GPT-5.3 Codex Spark
OpenAI
|
$0.025484 (rounded ~ $0.03) | ↑ 277.5% more | ↓ 52.9% cheaper |
| #17 |
GPT-5.3 Instant
OpenAI
|
$0.025484 (rounded ~ $0.03) | ↑ 277.5% more | ↓ 52.9% cheaper |
| #18 |
Grok 4.3
xAI
|
$0.028250 (rounded ~ $0.03) | ↑ 318.5% more | ↓ 47.7% cheaper |
| #19 |
Grok 4.20 Beta
xAI
|
$0.028250 (rounded ~ $0.03) | ↑ 318.5% more | ↓ 47.7% cheaper |
| #20 |
GPT-5.6 Terra
OpenAI
|
$0.030156 | ↑ 346.8% more | ↓ 44.2% cheaper |
| #21 |
Claude Sonnet 4.6
Anthropic
|
$0.032438 (rounded ~ $0.03) | ↑ 380.6% more | ↓ 40% cheaper |
| #22 |
Gemini 2.5 Pro
Google
|
$0.036406 (rounded ~ $0.04) | ↑ 439.4% more | ↓ 32.7% cheaper |
| #23 |
Gemini 3.1 Pro
Google
|
$0.048250 (rounded ~ $0.05) | ↑ 614.8% more | ↓ 10.8% cheaper |
| #24 |
Claude Opus 4.7
Anthropic
|
$0.054063 (rounded ~ $0.05) | ↑ 700.9% more | Same price |
| #25 |
Claude Opus 5
Anthropic
|
$0.054063 (rounded ~ $0.05) | ↑ 700.9% more | Same price |
| #26 |
Claude Opus 4.8
Anthropic
|
$0.054063 (rounded ~ $0.05) | ↑ 700.9% more | Same price |
| #27 |
Claude Opus 4.6
Anthropic
|
$0.054063 (rounded ~ $0.05) | ↑ 700.9% more | Same price |
| #28 |
GPT-5.4
OpenAI
|
$0.060313 | ↑ 793.5% more | ↑ 11.6% more |
| #29 |
GPT-5.4 Thinking
OpenAI
|
$0.060313 | ↑ 793.5% more | ↑ 11.6% more |
| #30 |
GPT-5.5 Instant
OpenAI
|
$0.060313 | ↑ 793.5% more | ↑ 11.6% more |
| #31 |
GPT-5.6 Sol
OpenAI
|
$0.060313 | ↑ 793.5% more | ↑ 11.6% more |
| #32 |
o3 Deep Research
OpenAI
|
$0.095625 (rounded ~ $0.10) | ↑ 1316.7% more | ↑ 76.9% more |
| #33 |
Claude Fable 5.1
Anthropic
|
$0.106719 (rounded ~ $0.11) | ↑ 1481% more | ↑ 97.4% more |
| #34 |
Claude Mythos 5.1
Anthropic
|
$0.106719 (rounded ~ $0.11) | ↑ 1481% more | ↑ 97.4% more |
| #35 |
Claude Fable 5
Anthropic
|
$0.108125 (rounded ~ $0.11) | ↑ 1501.9% more | ↑ 100% more |
| #36 |
Claude Mythos 5
Anthropic
|
$0.108125 (rounded ~ $0.11) | ↑ 1501.9% more | ↑ 100% more |
| #37 |
GPT-5.5
OpenAI
|
$0.120625 | ↑ 1687% more | ↑ 123.1% more |
| #38 |
o3 Pro
OpenAI
|
$0.191250 (rounded ~ $0.19) | ↑ 2733.3% more | ↑ 253.8% more |
| #39 |
GPT-6 Astra
OpenAI
|
$0.216250 (rounded ~ $0.22) | ↑ 3103.7% more | ↑ 300% more |
| #40 |
GPT-5.2 Pro
OpenAI
|
$0.305813 (rounded ~ $0.31) | ↑ 4430.6% more | ↑ 465.7% more |
| #41 |
GPT-5.2 Pro
OpenAI
|
$0.305813 (rounded ~ $0.31) | ↑ 4430.6% more | ↑ 465.7% more |
Mistral Small 3 Mistral AI
Gemini 3.1 Flash Lite Google
Mistral Large 3 Mistral AI
Gemini 3.5 Flash-Lite Google
Gemini 2.5 Flash Google
Gemini 3.8 Flash Google
GPT-5.4 mini OpenAI
o4-mini Deep Research OpenAI
o4-mini OpenAI
Claude Haiku 4.5 Anthropic
Gemini 3.1 Flash Google
GPT-5.6 Luna OpenAI
Gemini 3.6 Flash Google
Gemini 3.5 Flash Google
Claude Sonnet 5 Anthropic
GPT-5.3 Codex Spark OpenAI
GPT-5.3 Instant OpenAI
Grok 4.3 xAI
Grok 4.20 Beta xAI
GPT-5.6 Terra OpenAI
Claude Sonnet 4.6 Anthropic
Gemini 2.5 Pro Google
Gemini 3.1 Pro Google
Claude Opus 4.7 Anthropic
Claude Opus 5 Anthropic
Claude Opus 4.8 Anthropic
Claude Opus 4.6 Anthropic
GPT-5.4 OpenAI
GPT-5.4 Thinking OpenAI
GPT-5.5 Instant OpenAI
GPT-5.6 Sol OpenAI
o3 Deep Research OpenAI
Claude Fable 5.1 Anthropic
Claude Mythos 5.1 Anthropic
Claude Fable 5 Anthropic
Claude Mythos 5 Anthropic
GPT-5.5 OpenAI
o3 Pro OpenAI
GPT-6 Astra OpenAI
GPT-5.2 Pro OpenAI
GPT-5.2 Pro OpenAI
Running multiple autonomous agents on an engineering project.
Estimate the operational costs for 2026 enterprise AI deployments using latest tokenomics.