chiprook
← AI
AIOctober 6, 2026, 07:42

180B Darwin model runs on a laptop without GPU via 4-bit GGUF

The POCKET-Darwin-180B-GGUF build packs the 180-billion-parameter MoE model Darwin-180B-RSI into 111 GB and runs it on CPU: a 16-thread server CPU hits 18.4–21 tokens/s, while an RTX 5060 laptop with 32 GB RAM reaches 4.17 tokens/s. On MMLU-Pro the quantized build matched the original at 87.65%.

180B Darwin model runs on a laptop without GPU via 4-bit GGUF
#Darwin-180B-RSI#Llama.cpp#HuggingFace
Read next
AI

Swift 1.5 Qwen3.8-27B: GGUF variant built to stop overthinking

AI

Google TurboQuant compresses KV cache to 3 bits with no accuracy loss

AI

Viedraft's Darwin-180B-RSI Tops 10 Categories on Hugging Face Leaderboard

AI

Self-taught Darwin-180B-RSI tops Swiss law-exam benchmark LEXam