chiprook
← AI
AIOctober 10, 2026, 04:17

LightOnOCR 3 grounds extracted text back to page coordinates

LightOn released LightOnOCR 3 under Apache 2.0 in 0.8B, 1B and 4B sizes. Its grounding mode returns a bounding box for every extracted block, letting a chart value or table cell point back to the exact spot on the scanned page. Grounding adds roughly 25% more tokens, and the pinned vLLM setup breaks the 1B model if Transformers is upgraded manually.

LightOnOCR 3 grounds extracted text back to page coordinates
#LightOn#LightOnOCR
Read next
AI

LlamaIndex launches Extract v2.5 with accuracy and grounding gains

Policy

Aerospace industry pushes back on SpaceX-backed spectrum coordination overhaul

Gadgets

Amazon brings back Kindle page-turn buttons — but only in an $80 accessory

AI

LLM post-training digest: GRPO moves to OCR, DPO and GRPO get debugged