深潮TechFlow|Sep 21, 2026 15:24
[GLM-5.3-FlashX Officially Launches on B.AI, API and Web Chat Simultaneously Open]
According to Deep Tide TechFlow, on September 21, the B.AI platform announced the official launch of Z.AI's speed-optimized native multimodal large model GLM-5.3-FlashX, with API and Web Chat simultaneously open for access. This model is built on an efficient sparse architecture with 320B total parameters and 18B active parameters, boosting generation speed to a maximum of 200 Tokens/s while maintaining its original intelligence level—approximately 5 times faster than GLM-5.3-Flash. The model supports an ultra-long 1M Token context and comprehensive native video/file understanding, combined with interleaved tool invocation reasoning capabilities, specifically designed for high-frequency interactive programming and Agent workflows. Starting today, developers can directly access GLM-5.3-FlashX via the B.AI platform. Visit chat.b.ai/chat now to experience it firsthand.
Share To
Timeline
HotFlash
APP
X
Telegram
CopyLink