Zhipu launches GLM-5.3-FlashX with up to 200 tokens/s
Zhipu released the GLM-5.3-FlashX with speeds up to 200 tokens per second. The model is aimed at faster, smoother AI experiences for enterprise and developers, and its API is now online. This follows earlier exposure of the GLM-5.3-Flash under the Ox Alpha name.

Key facts
- 01GLM-5.3-FlashX speeds up to 200 tokens per second
- 02API is online
- 03GLM-5.3-Flash previously known as Ox Alpha
AI-generated from the sources below. Always check the originals.
How each country tells it
So far one country has covered this event. Perspectives appear when media in a second country report it.
Sources
Summaries are AI-generated from the linked sources and may contain errors; always check the originals. We summarise and link; we never republish articles. Photos come from openly licensed libraries, official publicity material and brand logos, credited to their sources. If you own an image and want it credited differently or removed, email info@coda.news and we will act promptly.
- 智谱 GLM-5.3-FlashX 模型上线,更快、更流畅 · IT之家
- 智谱推出GLM-5.3-FlashX · 36氪