2026-09-22 00:20
AI123
智谱上线更快速版GLM-5.3-Flash,推理速度最高200 tokens/s
智谱AI宣布,更快速的GLM-5.3-Flash模型正式上线,模型代码为glm-5.3-flashx,推理速度最高可达200 tokens/s。该模型面向所有API用户开放,Coding Plan用户可通过表单申请使用。
GLM系列是智谱AI推出的大语言模型系列,此次上线的更快速版本在命名和定价上均与现有GLM-5.3-Flash区分:新模型在Coding Plan和API上的定价均为GLM-5.3-Flash的2.5倍。
来源
- 2026-09-22 00:04 | X:RT Zixuan Li: Faster GLM-5.3-Flash is now live: up to 200 tokens/s. Model code: glm-5.3-flashx. Priced at 2.5× GLM-5.3-Flash on both the Coding Plan ...阅读原文
RT Zixuan Li Faster GLM-5.3-Flash is now live: up to 200 tokens/s. Model code: glm-5.3-flashx. Priced at 2.5× GLM-5.3-Flash on both the Coding Plan and API. Open to all API users. Coding Plan users can apply here: https://docs.google.com/forms/d/e/1FAIpQLSeDRfHS7i7zrZGsJheEADzzhCrHgsrXyZKjLtBPjvJIFN...