chiprook
← AI
AISeptember 17, 2026, 07:45

Nunchux AI Introduces VC-Attention: A Training-Free Low-Bit Attention Kernel That Speeds Up Video Diffusion Transformers

Nunchux AI released VC-Attention, a training-free low-bit attention kernel for video Diffusion Transformers. The method reduces value quantization error and speeds up softmax: on B200, attention in Wan2.2 runs 1.59 times faster, on RTX 5090 — 3.58 times faster.

Nunchux AI Introduces VC-Attention: A Training-Free Low-Bit Attention Kernel That Speeds Up Video Diffusion Transformers
#NunchuxAI#Wan2.2#MiniMax-H3#Nvidia
Read next
AI

Pentagon Review Links Palantir's Maven AI to Strike That Killed 120 Iranian Children

AI

David Pogue runs 125 tests on the new AI Siri in OS 27

AI

Viral Screenshots Claim ChatGPT Emailed the FBI From a User's Gmail Unprompted

AI

Meta's Muse AI Agent Tops U.S. iPhone Free-App Chart