Gemini 3.8 Flash vs Qwen3.8 Flash pricing
- For the Chatbot workload, Qwen3.8 Flash costs 6.8× less than Gemini 3.8 Flash.
- For the RAG workload, Qwen3.8 Flash costs 5.5× less than Gemini 3.8 Flash.
- For the Coding agent workload, Qwen3.8 Flash costs 5.8× less than Gemini 3.8 Flash.
Prices
| USD per 1M tokens | Gemini 3.8 Flash | Qwen3.8 Flash |
|---|---|---|
| Input | $0.75 | $0.15 |
| Output | $3.75 | $0.47 |
| Cache read | $0.075 | $0.016 |
| Cache write | Not published | $0.20 |
| Batch input | $0.375 | Not published |
| Batch output | $1.875 | Not published |
Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page Long-context and separately billed token prices, where shown, are list prices compiled by models.dev (checked Oct 10, 2026, 09:00 UTC).
Alibaba's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation
Monthly cost by workload
| Workload | Gemini 3.8 Flash | Qwen3.8 Flash |
|---|---|---|
| Chatbot1,000 in / 500 out tokens, 10,000 requests a day | $787.50/mo | $115.50/mo |
| RAG8,000 in / 500 out tokens, 5,000 requests a day | $1,181/mo | $215.25/mo |
| Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input | $540.00/mo | $92.40/mo |
Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.
Context and capabilities
| Gemini 3.8 Flash | Qwen3.8 Flash | |
|---|---|---|
| Context | 1,048,576 tokens | 1,000,000 tokens |
| Max output | 65,536 tokens | 131,072 tokens |
| Input | text, image, video, audio, pdf | text, image, video |
| Tool calling | Yes | Yes |
| Reasoning | Yes | Yes |
| Structured outputs | Yes | Yes |
Questions
- What is the price difference between Gemini 3.8 Flash and Qwen3.8 Flash?
- Gemini 3.8 Flash: $0.75 per 1M input tokens and $3.75 per 1M output tokens. Qwen3.8 Flash: $0.15 per 1M input tokens and $0.47 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $787.50 per month for Gemini 3.8 Flash and $115.50 for Qwen3.8 Flash.
- Which has the longer context window, Gemini 3.8 Flash or Qwen3.8 Flash?
- Gemini 3.8 Flash has the longer context window: up to 1,048,576 tokens, compared with 1,000,000 for Qwen3.8 Flash.
- Do Gemini 3.8 Flash and Qwen3.8 Flash have prompt caching and batch prices?
- Gemini 3.8 Flash: cached input $0.075 per 1M tokens; batch $0.375 per 1M input tokens and $1.875 per 1M output tokens. Qwen3.8 Flash: cached input $0.016 per 1M tokens; no batch price published.