B.AI Platform Unveils GLM-5.3-FlashX Multimodal AI Model
GLM-5.3-FlashX has launched on the B.AI platform, delivering high-speed coding and agent workflows through its native multimodal model.
The GLM-5.3-FlashX model is built on a 320B total/18B active parameter sparse architecture, allowing it to process up to 200 tokens per second with a 1 million context window.
This speed-optimized model is now live on B.AI API and Web Chat, enabling users to perform high-speed coding and agent workflows.