MiniMax-M3 vs Qwen3.8 Max pricing

Prices

MiniMax-M3 and Qwen3.8 Max prices, USD per 1M tokens
USD per 1M tokensMiniMax-M3Qwen3.8 Max
Input$0.30$2.00
Output$1.20$6.00
Cache read$0.06$0.25
Cache writeNot published$2.50
Batch inputNot publishedNot published
Batch outputNot publishedNot published

MiniMax's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation

Monthly cost by workload

WorkloadMiniMax-M3Qwen3.8 Max
Chatbot1,000 in / 500 out tokens, 10,000 requests a day$270.00/mo$1,500/mo
RAG8,000 in / 500 out tokens, 5,000 requests a day$450.00/mo$2,850/mo
Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input$234.00/mo$1,260/mo

Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.

Context and capabilities

MiniMax-M3Qwen3.8 Max
Context1,000,000 tokens1,000,000 tokens
Max output512,000 tokens131,072 tokens
Inputtext, image, videotext, image, video, pdf
Tool callingYesYes
ReasoningYesYes
Structured outputsNoYes

Questions

What is the price difference between MiniMax-M3 and Qwen3.8 Max?
MiniMax-M3: $0.30 per 1M input tokens and $1.20 per 1M output tokens. Qwen3.8 Max: $2.00 per 1M input tokens and $6.00 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $270.00 per month for MiniMax-M3 and $1,500 for Qwen3.8 Max.
Which has the longer context window, MiniMax-M3 or Qwen3.8 Max?
Both accept up to 1,000,000 tokens of context.
Do MiniMax-M3 and Qwen3.8 Max have prompt caching and batch prices?
MiniMax-M3: cached input $0.06 per 1M tokens; no batch price published. Qwen3.8 Max: cached input $0.25 per 1M tokens; no batch price published.

More comparisons