Gemini 3.8 Flash vs Grok Build 0.1 pricing
- For the Chatbot workload, Grok Build 0.1 costs 1.3× less than Gemini 3.8 Flash.
- For the RAG workload, Gemini 3.8 Flash costs 1.1× less than Grok Build 0.1.
- For the Coding agent workload, Gemini 3.8 Flash costs 1.2× less than Grok Build 0.1.
Prices
| USD per 1M tokens | Gemini 3.8 Flash | Grok Build 0.1 |
|---|---|---|
| Input | $0.75 | $1.00 |
| Output | $3.75 | $2.00 |
| Cache read | $0.075 | $0.20 |
| Cache write | Not published | Not published |
| Batch input | $0.375 | Not published |
| Batch output | $1.875 | Not published |
Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page Long-context and separately billed token prices, where shown, are list prices compiled by models.dev (checked Oct 10, 2026, 09:00 UTC).
xAI's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation
Monthly cost by workload
| Workload | Gemini 3.8 Flash | Grok Build 0.1 |
|---|---|---|
| Chatbot1,000 in / 500 out tokens, 10,000 requests a day | $787.50/mo | $600.00/mo |
| RAG8,000 in / 500 out tokens, 5,000 requests a day | $1,181/mo | $1,350/mo |
| Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input | $540.00/mo | $660.00/mo |
Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.
Context and capabilities
| Gemini 3.8 Flash | Grok Build 0.1 | |
|---|---|---|
| Context | 1,048,576 tokens | 256,000 tokens |
| Max output | 65,536 tokens | 256,000 tokens |
| Input | text, image, video, audio, pdf | text, image, pdf |
| Tool calling | Yes | Yes |
| Reasoning | Yes | Yes |
| Structured outputs | Yes | Yes |
Questions
- What is the price difference between Gemini 3.8 Flash and Grok Build 0.1?
- Gemini 3.8 Flash: $0.75 per 1M input tokens and $3.75 per 1M output tokens. Grok Build 0.1: $1.00 per 1M input tokens and $2.00 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $787.50 per month for Gemini 3.8 Flash and $600.00 for Grok Build 0.1.
- Which has the longer context window, Gemini 3.8 Flash or Grok Build 0.1?
- Gemini 3.8 Flash has the longer context window: up to 1,048,576 tokens, compared with 256,000 for Grok Build 0.1.
- Do Gemini 3.8 Flash and Grok Build 0.1 have prompt caching and batch prices?
- Gemini 3.8 Flash: cached input $0.075 per 1M tokens; batch $0.375 per 1M input tokens and $1.875 per 1M output tokens. Grok Build 0.1: cached input $0.20 per 1M tokens; no batch price published.