Gemini 3.8 Flash vs MiniMax-M3 pricing
- For the Chatbot workload, MiniMax-M3 costs 2.9× less than Gemini 3.8 Flash.
- For the RAG workload, MiniMax-M3 costs 2.6× less than Gemini 3.8 Flash.
- For the Coding agent workload, MiniMax-M3 costs 2.3× less than Gemini 3.8 Flash.
Prices
| USD per 1M tokens | Gemini 3.8 Flash | MiniMax-M3 |
|---|---|---|
| Input | $0.75 | $0.30 |
| Output | $3.75 | $1.20 |
| Cache read | $0.075 | $0.06 |
| Cache write | Not published | Not published |
| Batch input | $0.375 | Not published |
| Batch output | $1.875 | Not published |
Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page Long-context and separately billed token prices, where shown, are list prices compiled by models.dev (checked Oct 10, 2026, 09:00 UTC).
MiniMax's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation
Monthly cost by workload
| Workload | Gemini 3.8 Flash | MiniMax-M3 |
|---|---|---|
| Chatbot1,000 in / 500 out tokens, 10,000 requests a day | $787.50/mo | $270.00/mo |
| RAG8,000 in / 500 out tokens, 5,000 requests a day | $1,181/mo | $450.00/mo |
| Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input | $540.00/mo | $234.00/mo |
Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.
Context and capabilities
| Gemini 3.8 Flash | MiniMax-M3 | |
|---|---|---|
| Context | 1,048,576 tokens | 1,000,000 tokens |
| Max output | 65,536 tokens | 512,000 tokens |
| Input | text, image, video, audio, pdf | text, image, video |
| Tool calling | Yes | Yes |
| Reasoning | Yes | Yes |
| Structured outputs | Yes | No |
Questions
- What is the price difference between Gemini 3.8 Flash and MiniMax-M3?
- Gemini 3.8 Flash: $0.75 per 1M input tokens and $3.75 per 1M output tokens. MiniMax-M3: $0.30 per 1M input tokens and $1.20 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $787.50 per month for Gemini 3.8 Flash and $270.00 for MiniMax-M3.
- Which has the longer context window, Gemini 3.8 Flash or MiniMax-M3?
- Gemini 3.8 Flash has the longer context window: up to 1,048,576 tokens, compared with 1,000,000 for MiniMax-M3.
- Do Gemini 3.8 Flash and MiniMax-M3 have prompt caching and batch prices?
- Gemini 3.8 Flash: cached input $0.075 per 1M tokens; batch $0.375 per 1M input tokens and $1.875 per 1M output tokens. MiniMax-M3: cached input $0.06 per 1M tokens; no batch price published.