ModelInfo
简体中文
全部模型

GLM-5.3-FlashX

GLM-5.3-Flash 的高速变体,约 200 tokens/s。

Zhipu AIglm-5.3-flashx报告数据错误

规格

上下文
1.05M
最大输出
131.1K
输入价
$0.37/ 百万 Token
输出价
$1.25/ 百万 Token
发布日期
—

上下文

上下文
1.05M
最大输入
—
最大输出
131.1K

价格 · USD / 百万 Token

输入
$0.37
输出
$1.25
缓存读取
$0.075
缓存写入
—

输入输出

输入
文本图像视频文档
输出
文本

API

接口类型
chat

思考

支持思考
支持
思考等级
—

源信息

状态
可用
发布日期
—
知识截止
—

能力

已确认
toolsstructured_outputstreamingcaching
来源 · 官方文档核验 2026-10-10

调用方式

1 个 Provider

zhipumodel = glm-5.3-flashxreasoning = thinking.type

chat
POSThttps://api.z.ai/api/paas/v4/chat/completions

JSON

标准格式
glm-5.3-flashx.json
{
  "id": "glm-5.3-flashx",
  "object": "model",
  "created": null,
  "owned_by": "zhipu",
  "name": "GLM-5.3-FlashX",
  "api": {
    "types": [
      "chat"
    ]
  },
  "limits": {
    "context": 1048576,
    "input": null,
    "output": 131072
  },
  "modalities": {
    "input": [
      "text",
      "image",
      "video",
      "document"
    ],
    "output": [
      "text"
    ]
  },
  "reasoning": {
    "supported": true,
    "efforts": []
  },
  "pricing": {
    "currency": "USD",
    "unit": "1M_tokens",
    "input": 0.37,
    "output": 1.25,
    "cache_read": 0.075,
    "cache_write": null
  },
  "features": [
    "tools",
    "structured_output",
    "streaming",
    "caching"
  ],
  "info": {
    "status": "active",
    "release_date": null,
    "knowledge_cutoff": null,
    "description": "High-speed GLM-5.3-Flash variant at about 200 tokens/s.",
    "docs": "https://docs.z.ai/guides/vlm/glm-5.3-flash",
    "verified_at": "2026-10-10"
  }
}

官方来源

5