No profile yet.
Zhipu released the GLM-5.3-FlashX with speeds up to 200 tokens per second. The model is aimed at faster, smoother AI experiences for enterprise and developers, and its API is now online. This follows earlier exposure of the GLM-5.3-Flash under the Ox Alpha name.