Zhipu Launches GLM-5.3-FlashX: Up to 200 tokens/s, Domestic Chips Power the "Speed Edition" Flagship
9.18 On September 18, Zhipu officially launched GLM-5.3-FlashX, delivering inference speeds of up to 200 tokens/s for a faster, smoother experience for enterprises and developers. The API is now live with the model key GLM-5.3-FlashX.
Read article