llmburn.

← All models

MiniMax: MiniMax M3

Last Verified 2026-10-01 ✓ 3 independent sources agree context 1,048,576 JSON ↓

Pricing matrix USD per 1M tokens

FieldConsensusEvidence
Input $0.3 ✓ 3 independent sources 3 sources
Output $1.20 ✓ 3 independent sources 3 sources
Cache read $0.06 ✓ 3 independent sources 3 sources
Cache write (5m) Unverified(未验证,不填 0) not published
Cache write (1h) Unverified(未验证,不填 0) not published
Reasoning tokens Unverified(未验证,不填 0) not published
证据链:各源原值并列(audit trail,点击展开)
FieldSourceRaw value (USD/1M)IndependentObserved
input openrouter 0.3000 yes 2026-10-01 08:00
input litellm 0.3000 yes 2026-10-01 08:00
input litellm <span class="badge mute">mirror</span> 0.3000 no 2026-10-01 08:00
input modelsdev 0.3000 yes 2026-10-01 08:00
output openrouter 1.2000 yes 2026-10-01 08:00
output litellm 1.2000 yes 2026-10-01 08:00
output litellm <span class="badge mute">mirror</span> 1.2000 no 2026-10-01 08:00
output modelsdev 1.2000 yes 2026-10-01 08:00
cache_read openrouter 0.0600 yes 2026-10-01 08:00
cache_read litellm 0.0600 yes 2026-10-01 08:00
cache_read litellm <span class="badge mute">mirror</span> 0.0600 no 2026-10-01 08:00
cache_read modelsdev 0.0600 yes 2026-10-01 08:00

Long-context tier jumps ⚠

When prompt tokens cross a threshold, the entire request is re-priced at the higher tier — not just the excess. This is the #1 hidden cost on long-context workloads.

$0.17563K tokens512K跨过 512,000 tokens:$0.1548 / 请求(整请求重算)
Prompt ≥ tokensinputoutputcache read
0 (base)$0.3$1.20$0.06
512,000未验证未验证未验证

Run this model through the real-bill calculator →

Sources

This is an estimate based on public sticker prices, not your invoice. Machine-readable: /data/v1/pricing/minimax-minimax-m3.json