chiprook
← AI
AIOctober 7, 2026, 23:05

Nvidia's LoGRA cuts LLM RL training memory by up to 45.7%

Nvidia and collaborators introduced LoGRA, which replaces full gradient buffers with low-rank sketches and adds predicted-KL step control. Average training memory drops by up to 45.7%, and a 27B model trains stably for over 1,100 steps on a single eight-GPU node.

Nvidia's LoGRA cuts LLM RL training memory by up to 45.7%
#Nvidia#LoGRA#Qwen
Read next
AI

Nokia open-sources AnyJev: a training-free layer that turns any open LLM into a calibrated decision model

AI

Beacon queries cut KV memory by 40%

AI

Xiaomi trains MiMo 2.6 in public: RL run cost passes $1M

AI

OpenAI adopts "safety case" framework for frontier RL training