chiprook
← AI
AISeptember 24, 2026, 09:00

Xiaomi details HySparse2 for MiMo-V3: two-level KV sharing cuts 1M-token prefill FLOPs about 5x

Xiaomi's LLM-Core researchers released HySparse2, a hybrid sparse attention architecture with two-level KV sharing for long-horizon MiMo-V3-class agent workloads, in an arXiv paper. On matched 80B-A3B MoE models it lifts MRCR-v2 and RULER-v2 by 11.30 and 19.81 points over HySparse, and at 1M tokens cuts prefill FLOPs by 2.92x and 5.02x versus HySparse and Hybrid SWA.

Xiaomi details HySparse2 for MiMo-V3: two-level KV sharing cuts 1M-token prefill FLOPs about 5x
#Xiaomi#MiMo#HySparse2
Read next
AI

OpenAI AI Tried to Breach Four Other Targets Without Prompting

AI

Google Nears Release of Flagship Gemini 4 AI Model

AI

OpenAI launches benchmark for mental health AI

AI

TypeSafe Jev: Level With Mid-Price LLMs, Behind the Frontier