[!NOTE] π LightOnOCR 2 is now available and state of the art on OlmOCR bench , with new image detection variants! Check it out here: lightonai/LightOnOCR 2 1B LightOnOCR 1B 1025 Full BF16 version of the model. We recommend this variant for inference and further fine tuning. LightOnOCR 1B is a compact, end to end visionβlanguage model for Optical Character Recognition (OCR) and document understanding. It achieves state of the art accuracy in its weight class while being several times faster and cheaper than larger general purpose VLMs. π Paper π Read the full blog post π Try the demo π Finetuning notebook Highlights β‘ Speed: 5Γ faster than dots.ocr, 2Γ faster than PaddleOCR VL 0.9B, 1.73Γ faster than DeepSeekOCR πΈ Efficiency: Processes 5.71 pages/s on a single H100 (~493k pages/day) for All benchmarks evaluated using vLLM on the Olmo Bench. Installation VLLM [2025/11/24] π LightOnOCR is now officially supported in vLLM v0.11.1 π Start Server PDF Inference Transformers Note: LightOnOCR 2 requires transformers installed from source (not yet in a stable release). Rendering and Preprocessing Tips Render PDFs to PNG or JPEG at a target longest dimension of 1540px Maintain aspectβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy