ModelInfo
English
All models

GLM-4-FlashX-250414

Fast, high-concurrency enhanced version of the free GLM-4-Flash with long-context and multilingual support.

Zhipu AIglm-4-flashx-250414Report incorrect data

Specs

Context
131.1K
Max output
16.4K
Input price
¥0.1/ 1M tokens
Output price
¥0.1/ 1M tokens
Released
—

Context

Context window
131.1K
Max input
—
Max output
16.4K

Pricing · CNY / 1M tokens

Input
¥0.1
Output
¥0.1
Cache read
¥0.05
Cache write
—

Modalities

Input
Text
Output
Text

API

API types
chat

Reasoning

Not published

Info

Status
Active
Released
—
Knowledge cutoff
—

Features

Confirmed
toolsstructured_outputstreamingcaching
Source · official docsVerified 2026-10-11

How to call

1 provider

zhipumodel = glm-4-flashx-250414

chat
POSThttps://open.bigmodel.cn/api/paas/v4/chat/completions

JSON

Standard format
glm-4-flashx-250414.json
{
  "id": "glm-4-flashx-250414",
  "object": "model",
  "created": null,
  "owned_by": "zhipu",
  "name": "GLM-4-FlashX-250414",
  "api": {
    "types": [
      "chat"
    ]
  },
  "limits": {
    "context": 131072,
    "input": null,
    "output": 16384
  },
  "modalities": {
    "input": [
      "text"
    ],
    "output": [
      "text"
    ]
  },
  "reasoning": {
    "supported": null,
    "efforts": []
  },
  "pricing": {
    "currency": "CNY",
    "unit": "1M_tokens",
    "input": 0.1,
    "output": 0.1,
    "cache_read": 0.05,
    "cache_write": null
  },
  "features": [
    "tools",
    "structured_output",
    "streaming",
    "caching"
  ],
  "info": {
    "status": "active",
    "release_date": null,
    "knowledge_cutoff": null,
    "description": "Fast, high-concurrency enhanced version of the free GLM-4-Flash with long-context and multilingual support.",
    "docs": "https://docs.bigmodel.cn/cn/guide/models/text/glm-4",
    "verified_at": "2026-10-11"
  }
}

Official sources

5