Claude Opus 5.5 vs Gemini 3.5 Flash Lite pricing

Prices

Claude Opus 5.5 and Gemini 3.5 Flash Lite prices, USD per 1M tokens
USD per 1M tokensClaude Opus 5.5Gemini 3.5 Flash Lite
Input$4.00$0.30
Output$20.00$2.50
Cache read$0.20$0.03
Cache write$5.00Not published
Batch input$2.00$0.15
Batch output$10.00$1.25

Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page

Monthly cost by workload

WorkloadClaude Opus 5.5Gemini 3.5 Flash Lite
Chatbot1,000 in / 500 out tokens, 10,000 requests a day$4,200/mo$465.00/mo
RAG8,000 in / 500 out tokens, 5,000 requests a day$6,300/mo$547.50/mo
Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input$2,640/mo$276.00/mo

Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.

Context and capabilities

Claude Opus 5.5Gemini 3.5 Flash Lite
Context1,000,000 tokens1,048,576 tokens
Max output128,000 tokens65,536 tokens
Inputtext, image, pdftext, image, video, audio, pdf
Tool callingYesYes
ReasoningYesYes
Structured outputsYesYes

Questions

What is the price difference between Claude Opus 5.5 and Gemini 3.5 Flash Lite?
Claude Opus 5.5: $4.00 per 1M input tokens and $20.00 per 1M output tokens. Gemini 3.5 Flash Lite: $0.30 per 1M input tokens and $2.50 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $4,200 per month for Claude Opus 5.5 and $465.00 for Gemini 3.5 Flash Lite.
Which has the longer context window, Claude Opus 5.5 or Gemini 3.5 Flash Lite?
Gemini 3.5 Flash Lite has the longer context window: up to 1,048,576 tokens, compared with 1,000,000 for Claude Opus 5.5.
Do Claude Opus 5.5 and Gemini 3.5 Flash Lite have prompt caching and batch prices?
Claude Opus 5.5: cached input $0.20 per 1M tokens; batch $2.00 per 1M input tokens and $10.00 per 1M output tokens. Gemini 3.5 Flash Lite: cached input $0.03 per 1M tokens; batch $0.15 per 1M input tokens and $1.25 per 1M output tokens.

More comparisons