chiprook
← AI
AISeptember 18, 2026, 14:51

Zhipu Opens GLM-5.3-FlashX Near 200 Tokens/s on ~100k Domestic Accelerators

Zhipu AI opened access to GLM-5.3-FlashX via API and testing center: peak generation about 200 tokens per second on approximately 100,000 Chinese AI accelerators. The base model GLM-5.3-Flash is a 320B-parameter MoE (18B active) with 1M token context, open-sourced on August 26.

Zhipu Opens GLM-5.3-FlashX Near 200 Tokens/s on ~100k Domestic Accelerators
#Zhipu#GLM
Read next
AI

OpenAI builds features to counter Grok Bot, weighs answer to Meta's agent

AI

OpenAI and Anthropic Negotiated Mutual AI Stress-Testing Deal

AI

AI tag team Jev and Astra beat Minecraft in 8m43s for under $1

AI

OpenAI details six real model misalignment incidents