← Back to live feed · 1 stories across 1 day

Monday, Sep 21, 2026

1 story
1
GLM 5.3 FlashX Hits 200 Tokens Per Second at 2.5X Price
topics 🤖 AI tags AIAI ModelsAI Releases keywords Zixuan Li

Zixuan Li announced the launch of a high-speed variant of the GLM Flash model for developers and enterprise users. The glm-5.3-flashx model provides throughput of up to 200 tokens/s, carrying a price 2.5x higher than the base model on both the API and Coding Plan. General API users have immediate access to the new version, though those seeking access via the Coding Plan must submit an application.

The model is now available in the Hermes Agent through the Nous Portal and OpenRouter. This iteration focuses on delivery speed for large language model responses, targeting users who require faster token output than standard Flash versions provide for throughput-heavy AI tasks.

Earlier version from Sunday, Sep 20
OpenRouter and Nous Portal Launch Zhipu AI GLM-5.3 FlashX at 200 Tokens Per Second
6 tweets • 6 sources
See all 5 tweets →