chiprook
← AI
AIOctober 9, 2026, 04:56

MLPerf Inference v6.1 adds end-to-end RAG and desktop agent benchmarks

MLCommons released MLPerf Inference v6.1 with two new tests: an End-to-End RAG benchmark for the full retrieval pipeline and an Edge Agentic benchmark replaying 1,007 turns of a coding agent on a desktop. A record 30 organizations submitted 120 systems and 486 results; DeepSeek-R1 is 5.7x faster than a year ago.

MLPerf Inference v6.1 adds end-to-end RAG and desktop agent benchmarks
#Nvidia#AMD#MLCommons
Read next
Hardware

MLPerf Inference v6.1: 5.7x per-accelerator gains, 512-GPU run, Vera Rubin results

Hardware

AMD benchmarks 512 MI355X GPU cluster in MLPerf 6.1 while NVIDIA retains strong position with Blackwell & first Vera Rubin submit, Intel Arc Pro & Xeon submitted too

AI

Horizon Robotics releases HSD V2.1 with end-to-end reversing

AI

laya-compactor cuts RAG context tokens by 70% with no answer-quality loss