DeepSeek V4 (Engram) DeepSeek 1000000
💰 Total Cost Calculation (from Plugin)
Output: $0.087000 (rounded ~ $0.09)
Output: $0.087000 (rounded ~ $0.09)
Unit: $0.000000
Fees: $0.000000
Detailed Cost Analysis (from Plugin)
For 500,000 input tokens and 100,000 output tokens:
- Input Cost: $0.217500 (rounded ~ $0.22)
- Output Cost: $0.087000 (rounded ~ $0.09)
- Total Cost: $0.112665 (rounded ~ $0.11)
- Cost per 1K tokens: $0.000188
- Tokens per dollar: 5,325,523 tokens
- Context Window: 1000000 tokens
Speed & Performance Analysis
With a processing speed of 300 tokens per second and 180ms time to first token:
- Processing Time: 34 minutes, 40.18 seconds
- Latency: 180 milliseconds to first token
- Base Throughput: 300 tokens/second
- Effective Throughput: 288 tokens/second (temperature-adjusted)
Best Use Cases
Want this applied to YOUR actual stack?
This calculator shows the math for DeepSeek V4 (Engram). Your decision needs more — current infrastructure, compliance requirements, actual workload patterns, volume tiers — that change which model is right for you.
Get a $39 personalized AI Architecture Audit. PDF tailored to your stack, delivered in under 60 seconds. 7-day no-questions-asked refund.
Get my instant AI audit — $39 →Claude Sonnet 4.6 Anthropic 1000000
💰 Total Cost Calculation (from Plugin)
Output: $0.375000 (rounded ~ $0.38)
Output: $0.375000 (rounded ~ $0.38)
Unit: $0.000000
Fees: $0.000000
Advanced Cost Breakdown (from Plugin)
Detailed Cost Analysis (from Plugin)
For 500,000 input tokens and 100,000 output tokens:
- Input Cost: $0.375000 (rounded ~ $0.38)
- Output Cost: $0.375000 (rounded ~ $0.38)
- Total Cost: $0.446250 (rounded ~ $0.45)
- Cost per 1K tokens: $0.000744
- Tokens per dollar: 1,344,538 tokens
- Context Window: 1000000 tokens
Speed & Performance Analysis
With a processing speed of 450 tokens per second and 200ms time to first token:
- Processing Time: 23 minutes, 6.85 seconds
- Latency: 200 milliseconds to first token
- Base Throughput: 450 tokens/second
- Effective Throughput: 433 tokens/second (temperature-adjusted)
Best Use Cases
Want this applied to YOUR actual stack?
This calculator shows the math for Claude Sonnet 4.6. Your decision needs more — current infrastructure, compliance requirements, actual workload patterns, volume tiers — that change which model is right for you.
Get a $39 personalized AI Architecture Audit. PDF tailored to your stack, delivered in under 60 seconds. 7-day no-questions-asked refund.
Get my instant AI audit — $39 →✨ Market Recommendations AI Model Registry
← Back to DeepSeek V4 (Engram)| Rank | AI Model & Provider | Total Cost | vs DeepSeek V4 (Engram) | vs Claude Sonnet 4.6 |
|---|---|---|---|---|
| 🏆 |
Gemini 3.1 Flash Lite
Google
|
$0.043438 (rounded ~ $0.04) Best Value | ↓ 61.4% cheaper | ↓ 90.3% cheaper |
| 🥈 |
Gemini 3.5 Flash-Lite
Google
|
$0.069625 | ↓ 38.2% cheaper | ↓ 84.4% cheaper |
| 🥉 |
Gemini 2.5 Flash
Google
|
$0.069625 | ↓ 38.2% cheaper | ↓ 84.4% cheaper |
| #4 |
Gemini 3.8 Flash
Google
|
$0.111563 (rounded ~ $0.11) | ↓ 1% cheaper | ↓ 75% cheaper |
| #5 |
GPT-5.6 Luna
OpenAI
|
$0.173750 (rounded ~ $0.17) | ↑ 54.2% more | ↓ 61.1% cheaper |
| #6 |
Gemini 3.6 Flash
Google
|
$0.223125 (rounded ~ $0.22) | ↑ 98% more | ↓ 50% cheaper |
| #7 |
Gemini 3.5 Flash
Google
|
$0.260625 | ↑ 131.3% more | ↓ 41.6% cheaper |
| #8 |
Claude Sonnet 5
Anthropic
|
$0.297500 (rounded ~ $0.30) | ↑ 164.1% more | ↓ 33.3% cheaper |
| #9 |
Gemini 3.1 Flash
Google
|
$0.347500 (rounded ~ $0.35) | ↑ 208.4% more | ↓ 22.1% cheaper |
| #10 |
GPT-5.6 Terra
OpenAI
|
$0.434375 (rounded ~ $0.43) | ↑ 285.5% more | ↓ 2.7% cheaper |
| #11 |
Claude Sonnet 4.6
Anthropic
|
$0.446250 (rounded ~ $0.45) | ↑ 296.1% more | Same price |
| #12 |
Grok 4.3
xAI
|
$0.590000 | ↑ 423.7% more | ↑ 32.2% more |
| #13 |
Grok 4.20 Beta
xAI
|
$0.590000 | ↑ 423.7% more | ↑ 32.2% more |
| #14 |
Claude Opus 4.7
Anthropic
|
$0.743750 (rounded ~ $0.74) | ↑ 560.1% more | ↑ 66.7% more |
| #15 |
Claude Opus 5
Anthropic
|
$0.743750 (rounded ~ $0.74) | ↑ 560.1% more | ↑ 66.7% more |
| #16 |
Claude Opus 4.8
Anthropic
|
$0.743750 (rounded ~ $0.74) | ↑ 560.1% more | ↑ 66.7% more |
| #17 |
Claude Opus 4.6
Anthropic
|
$0.743750 (rounded ~ $0.74) | ↑ 560.1% more | ↑ 66.7% more |
| #18 |
Gemini 2.5 Pro
Google
|
$0.868750 (rounded ~ $0.87) | ↑ 671.1% more | ↑ 94.7% more |
| #19 |
GPT-5.6 Sol
OpenAI
|
$0.868750 (rounded ~ $0.87) | ↑ 671.1% more | ↑ 94.7% more |
| #20 |
Gemini 3.1 Pro
Google
|
$1.090000 | ↑ 867.5% more | ↑ 144.3% more |
| #21 |
GPT-5.4
OpenAI
|
$1.362500 (rounded ~ $1.36) | ↑ 1109.3% more | ↑ 205.3% more |
| #22 |
GPT-5.4 Thinking
OpenAI
|
$1.362500 (rounded ~ $1.36) | ↑ 1109.3% more | ↑ 205.3% more |
| #23 |
Claude Fable 5.1
Anthropic
|
$1.403125 (rounded ~ $1.40) | ↑ 1145.4% more | ↑ 214.4% more |
| #24 |
Claude Mythos 5.1
Anthropic
|
$1.403125 (rounded ~ $1.40) | ↑ 1145.4% more | ↑ 214.4% more |
| #25 |
Claude Fable 5
Anthropic
|
$1.487500 (rounded ~ $1.49) | ↑ 1220.3% more | ↑ 233.3% more |
| #26 |
Claude Mythos 5
Anthropic
|
$1.487500 (rounded ~ $1.49) | ↑ 1220.3% more | ↑ 233.3% more |
| #27 |
GPT-5.5
OpenAI
|
$2.725000 (rounded ~ $2.73) | ↑ 2318.7% more | ↑ 510.6% more |
| #28 |
GPT-6 Astra
OpenAI
|
$5.950000 | ↑ 5181.1% more | ↑ 1233.3% more |
| #29 |
GPT-6 Astra
OpenAI
|
$5.950000 | ↑ 5181.1% more | ↑ 1233.3% more |
Gemini 3.1 Flash Lite Google
Gemini 3.5 Flash-Lite Google
Gemini 2.5 Flash Google
Gemini 3.8 Flash Google
GPT-5.6 Luna OpenAI
Gemini 3.6 Flash Google
Gemini 3.5 Flash Google
Claude Sonnet 5 Anthropic
Gemini 3.1 Flash Google
GPT-5.6 Terra OpenAI
Claude Sonnet 4.6 Anthropic
Grok 4.3 xAI
Grok 4.20 Beta xAI
Claude Opus 4.7 Anthropic
Claude Opus 5 Anthropic
Claude Opus 4.8 Anthropic
Claude Opus 4.6 Anthropic
Gemini 2.5 Pro Google
GPT-5.6 Sol OpenAI
Gemini 3.1 Pro Google
GPT-5.4 OpenAI
GPT-5.4 Thinking OpenAI
Claude Fable 5.1 Anthropic
Claude Mythos 5.1 Anthropic
Claude Fable 5 Anthropic
Claude Mythos 5 Anthropic
GPT-5.5 OpenAI
GPT-6 Astra OpenAI
GPT-6 Astra OpenAI
Performance per Dollar
DeepSeek V4 (Engram) provides 1M context with ‘Engram Memory’ for just $0.27/1M input tokens. Claude Sonnet 4.6 is the industry favorite for ‘Agentic Coding’ reliability but costs $3.00/1M. DeepSeek V4 is the choice for high-volume automated refactoring, while Sonnet 4.6 remains the gold standard for complex system architecture.