LightOnOCR 2 1B base Base model for fine tuning. This is the pre RLVR checkpoint with strong OCR capabilities, ideal as a starting point for domain adaptation and custom fine tuning. About LightOnOCR 2 LightOnOCR 2 is an efficient end to end 1B parameter vision language model for converting documents (PDFs, scans, images) into clean, naturally ordered text without relying on brittle pipelines. This second version is trained on a larger and higher quality corpus with stronger French, arXiv, and scan coverage, improved LaTeX handling, and cleaner normalization. LightOnOCR 2 achieves state of the art performance on OlmOCR Bench while being ~9× smaller and significantly faster than competing approaches. Highlights ⚡ Speed: 3.3× faster than Chandra OCR, 1.7× faster than OlmOCR, 5× faster than dots.ocr, 2× faster than PaddleOCR VL 0.9B, 1.73× faster than DeepSeekOCR 💸 Efficiency: Processes 5.71 pages/s on a single H100 (~493k pages/day) for See the paper for full benchmark details and methodology. Usage with Transformers Note: LightOnOCR 2 requires transformers installed from source (not yet in a stable release). Usage with vLLM Rendering and Preprocessing Tips Render PDFs to PNG or JPE…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy