chiprook
← AI
AIOctober 6, 2026, 07:45

2026 open-weight MoE models activate just 1-9% of their weights

In 2026, open-weight sparse Mixture-of-Experts models activate only 1-9% of their weights per token: DeepSeek V4.1-Flash carries 552B total weights but activates ~8B, Kimi K3 carries 2.8T and activates ~104B, and GLM-5.2 runs ~744B total with ~40B active. Compute scales with active parameters while memory and loading scale with total parameters, diverging by one to two orders of magnitude.

2026 open-weight MoE models activate just 1-9% of their weights
#DeepSeek#Kimi#GLM#Qwen
Read next
AI

NVIDIA Build offers free API access to Kimi K3, DeepSeek and GLM 5.3

AI

GLM-5.3-Flash: 320B Open-Weight Model With 18B Active and 1M-Token Context

AI

Open-source coding LLMs compared: GLM-5.3-Flash, Qwen3.8-Flash-Next, DeepSeek V4 Flash

AI

Xiaomi open-sources MiMo-V2.6: 46 points and a $6 subscription