GLM-4-FlashX-250414
Fast, high-concurrency enhanced version of the free GLM-4-Flash with long-context and multilingual support.
Specs
- Context
- 131.1K
- Max output
- 16.4K
- Input price
- ¥0.1/ 1M tokens
- Output price
- ¥0.1/ 1M tokens
- Released
- —
Context
- Context window
- 131.1K
- Max input
- —
- Max output
- 16.4K
Pricing · CNY / 1M tokens
- Input
- ¥0.1
- Output
- ¥0.1
- Cache read
- ¥0.05
- Cache write
- —
Modalities
- Input
- Text
- Output
- Text
API
- API types
Reasoning
- Not published
Info
- Status
- Active
- Released
- —
- Knowledge cutoff
- —
Features
- Confirmed
How to call
1 provider
zhipumodel = glm-4-flashx-250414
chat
POSThttps://open.bigmodel.cn/api/paas/v4/chat/completions
JSON
Standard format
{
"id": "glm-4-flashx-250414",
"object": "model",
"created": null,
"owned_by": "zhipu",
"name": "GLM-4-FlashX-250414",
"api": {
"types": [
"chat"
]
},
"limits": {
"context": 131072,
"input": null,
"output": 16384
},
"modalities": {
"input": [
"text"
],
"output": [
"text"
]
},
"reasoning": {
"supported": null,
"efforts": []
},
"pricing": {
"currency": "CNY",
"unit": "1M_tokens",
"input": 0.1,
"output": 0.1,
"cache_read": 0.05,
"cache_write": null
},
"features": [
"tools",
"structured_output",
"streaming",
"caching"
],
"info": {
"status": "active",
"release_date": null,
"knowledge_cutoff": null,
"description": "Fast, high-concurrency enhanced version of the free GLM-4-Flash with long-context and multilingual support.",
"docs": "https://docs.bigmodel.cn/cn/guide/models/text/glm-4",
"verified_at": "2026-10-11"
}
}Official sources
5