chiprook
← AI
AIOctober 6, 2026, 04:54

Inception Labs launches Mercury 2.5, a diffusion LLM at 1,107 tokens/sec

Inception Labs announced Mercury 2.5, billed as the largest diffusion language model ever trained, running at 1,107 tokens per second on widely available Nvidia GPUs. It claims a 40% intelligence gain over Mercury 2, matches cost-optimized frontier tier quality, and ships via an OpenAI-compatible API with a 260K-token context window.

Inception Labs launches Mercury 2.5, a diffusion LLM at 1,107 tokens/sec
#InceptionLabs#Mercury2.5#Nvidia
Read next
AI

Google Research Introduces Retrieve-for-Train (R4T): RL-Compiled Diffusion Retriever for 12× to 20× Faster Query Fan-Out

AI

Nunchux AI Introduces VC-Attention: A Training-Free Low-Bit Attention Kernel That Speeds Up Video Diffusion Transformers

Science

Mercury Shrank More Than Thought, Losing Up to 14.5 Miles in Diameter

Science

After 8 Years in Space, BepiColombo Begins Mercury Orbit Insertion