chiprook
← AI
AISeptember 20, 2026, 10:44

Qwen3.8-Flash-Next vs Qwen3.8-27B: 125B Parameters, 6B Active — What the Qwen4 Preview Changes

Alibaba introduced the open MoE model Qwen3.8-Flash-Next: 125 billion parameters with 6 billion active per token and a 51 billion N-gram embedding table. Context is 262,144 tokens extendable to 1 million, claimed score of 62.5 on SWE-bench Pro, and training cost about 1/9 of Qwen3.7-Plus.

Qwen3.8-Flash-Next vs Qwen3.8-27B: 125B Parameters, 6B Active — What the Qwen4 Preview Changes
#Alibaba#Qwen
Read next
AI

AI robot arms attempted harmful tasks 97% of the time without jailbreaks

AI

DeepSeek to adopt Huawei chips for model training in Q4

AI

Nvidia's free AI model Nemotron 4 could pull the UAE closer to the U.S.

AI

Anthropic Sets Up Wet Lab for Biology Research