chiprook
← AI
AISeptember 25, 2026, 01:58

BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens

BottleCap AI has released ThinkingCap-Qwen3.8-27B, a fine-tune of Qwen3.8-27B that cuts thinking tokens by an average of 37.2% across 12 benchmarks while macro-average accuracy drops just 0.86pp, from 86.65% to 85.79%. It is a drop-in replacement for Qwen3.8-27B on vLLM or SGLang, with FP8, NVFP4, GGUF and MLX builds.

BottleCap AI Releases ThinkingCap-Qwen3.8-27B: 37.2% Fewer Thinking Tokens
#BottleCapAI#Qwen
Read next
AI

Okta turns Dex AI agent into a customer-zero proving ground

AI

Kalshi admits it used AI to turn a YouTuber's video into an ad, changing his race

AI

Microsoft AI Chief Suleyman Calls for a 'Red Line' on AI

AI

ChatGPT drives 95.1% of AI referral traffic to websites, BrightEdge says