chiprook
← AI
AIOctober 3, 2026, 14:45

vLLM ignores LoRA rank_pattern and alpha_pattern, mis-scaling adapters

vLLM 0.30.0 reads only r and lora_alpha from a PEFT LoRA adapter_config.json, silently dropping rank_pattern and alpha_pattern, so modules with per-module scaling are served at the wrong scale. In a Qwen3 test the mean prompt log-probability gap versus PEFT rose from 0.0003 to 0.0211; a per-module fix (#59801) is open but unmerged.

vLLM ignores LoRA rank_pattern and alpha_pattern, mis-scaling adapters
#VLLM#PEFT#Qwen3#OLMoE
Read next
AI

Alpha School students were monitored by offshore call center staff, not AI

AI

Union Alpha Is No Longer a Mystery Model and Stopped Being Free Yesterday

AI

Ai2 Releases Olmo-core 3, Open Training Stack for Trillion-Parameter MoEs

AI

Ai2 open-sources AstaBrief 8B for fast scientific report generation