GLM 5.3 Flash (batch)

GLM 5.3 Flash · 批量推理档(Batch)

异步批量调用档:以折扣价处理离线批量请求,吞吐高、延迟不敏感,适合大规模非实时任务。

适用场景:大规模离线推理、成本敏感批处理、跑批与回填

本页为 GLM 5.3 Flash 的批量推理档,与其实时标准档共享底层模型能力,主要差异在计费模式与调用方式。标准实时档详见 GLM 5.3 Flash 标准实时档

厂商:智谱 AI

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

关键参数

Token 定价

输入输出模态

文本、图像、视频

典型用例

什么时候选这个模型

基于公开定价与能力字段派生的选型信号(非实测benchmark),最终选型请结合官方文档与你的评测。

相关问题

同厂商其他模型

查看完整信息:GLM 5.3 Flash (batch)