GLM-5.3-FlashX vs GPT-6 Astra pricing

Prices

GLM-5.3-FlashX and GPT-6 Astra prices, USD per 1M tokens
USD per 1M tokensGLM-5.3-FlashXGPT-6 Astra
Input$0.37$10.00
Output$1.25$50.00
Cache read$0.075$1.00
Cache writeFree$12.50
Batch inputNot published$5.00
Batch outputNot published$25.00

Z.AI's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation

Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page

Monthly cost by workload

WorkloadGLM-5.3-FlashXGPT-6 Astra
Chatbot1,000 in / 500 out tokens, 10,000 requests a day$298.50/mo$10,500/mo
RAG8,000 in / 500 out tokens, 5,000 requests a day$537.75/mo$15,750/mo
Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input$276.00/mo$7,200/mo

Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.

Context and capabilities

GLM-5.3-FlashXGPT-6 Astra
Context1,000,000 tokens1,050,000 tokens
Max output131,072 tokens128,000 tokens
Inputtext, image, video, pdftext, image, pdf
Tool callingYesYes
ReasoningYesYes
Structured outputsYesYes

Questions

What is the price difference between GLM-5.3-FlashX and GPT-6 Astra?
GLM-5.3-FlashX: $0.37 per 1M input tokens and $1.25 per 1M output tokens. GPT-6 Astra: $10.00 per 1M input tokens and $50.00 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $298.50 per month for GLM-5.3-FlashX and $10,500 for GPT-6 Astra.
Which has the longer context window, GLM-5.3-FlashX or GPT-6 Astra?
GPT-6 Astra has the longer context window: up to 1,050,000 tokens, compared with 1,000,000 for GLM-5.3-FlashX.
Do GLM-5.3-FlashX and GPT-6 Astra have prompt caching and batch prices?
GLM-5.3-FlashX: cached input $0.075 per 1M tokens; no batch price published. GPT-6 Astra: cached input $1.00 per 1M tokens; batch $5.00 per 1M input tokens and $25.00 per 1M output tokens.

More comparisons