GLM-5.3-FlashX vs GPT-6.1 Sol pricing
- For the Chatbot workload, GLM-5.3-FlashX costs 7.0× less than GPT-6.1 Sol.
- For the RAG workload, GLM-5.3-FlashX costs 5.9× less than GPT-6.1 Sol.
- For the Coding agent workload, GLM-5.3-FlashX costs 4.8× less than GPT-6.1 Sol.
Prices
| USD per 1M tokens | GLM-5.3-FlashX | GPT-6.1 Sol |
|---|---|---|
| Input | $0.37 | $2.00 |
| Output | $1.25 | $10.00 |
| Cache read | $0.075 | $0.10 |
| Cache write | Free | $2.50 |
| Batch input | Not published | $1.00 |
| Batch output | Not published | $5.00 |
Z.AI's list price for its own API, as compiled by models.dev and checked Oct 10, 2026, 09:00 UTC. Prices on cloud platforms and resellers can differ. Provider documentation
Verified against the provider's official pricing page on Oct 9, 2026. Official pricing page
Monthly cost by workload
| Workload | GLM-5.3-FlashX | GPT-6.1 Sol |
|---|---|---|
| Chatbot1,000 in / 500 out tokens, 10,000 requests a day | $298.50/mo | $2,100/mo |
| RAG8,000 in / 500 out tokens, 5,000 requests a day | $537.75/mo | $3,150/mo |
| Coding agent50,000 in / 2,000 out tokens, 1,000 requests a day, 80% cached input | $276.00/mo | $1,320/mo |
Cost = input tokens × input price + output tokens × output price, 30 days a month. Cached input uses the cache read price. Cache write fees are not included.
Context and capabilities
| GLM-5.3-FlashX | GPT-6.1 Sol | |
|---|---|---|
| Context | 1,000,000 tokens | 1,050,000 tokens |
| Max output | 131,072 tokens | 128,000 tokens |
| Input | text, image, video, pdf | text, image, pdf |
| Tool calling | Yes | Yes |
| Reasoning | Yes | Yes |
| Structured outputs | Yes | Yes |
Questions
- What is the price difference between GLM-5.3-FlashX and GPT-6.1 Sol?
- GLM-5.3-FlashX: $0.37 per 1M input tokens and $1.25 per 1M output tokens. GPT-6.1 Sol: $2.00 per 1M input tokens and $10.00 per 1M output tokens. For the Chatbot workload (1,000 in / 500 out tokens, 10,000 requests a day), that is about $298.50 per month for GLM-5.3-FlashX and $2,100 for GPT-6.1 Sol.
- Which has the longer context window, GLM-5.3-FlashX or GPT-6.1 Sol?
- GPT-6.1 Sol has the longer context window: up to 1,050,000 tokens, compared with 1,000,000 for GLM-5.3-FlashX.
- Do GLM-5.3-FlashX and GPT-6.1 Sol have prompt caching and batch prices?
- GLM-5.3-FlashX: cached input $0.075 per 1M tokens; no batch price published. GPT-6.1 Sol: cached input $0.10 per 1M tokens; batch $1.00 per 1M input tokens and $5.00 per 1M output tokens.