olmOCR 2 7B 1025 FP8 Quantized to FP8 Version of olmOCR 2 7B 1025, using llmcompressor. This is a release of the olmOCR model that's fine tuned from Qwen2.5 VL 7B Instruct using the olmOCR mix 1025 dataset. It has been additionally fine tuned using GRPO RL training to boost its performance at math equations, tables, and other tricky OCR cases. Quick links: 📃 Paper 🤗 SFT Dataset 🤗 RL Dataset 🛠️ Code 🎮 Demo The best way to use this model is via the olmOCR toolkit. The toolkit comes with an efficient inference setup via VLLM that can handle millions of documents at scale. olmOCR Bench Scores This model scores the following scores on olmOCR bench when used with the olmOCR toolkit toolkit which automatically renders, rotates, and retries pages as needed. Model ArXiv Old Scans Math Tables Old Scans Headers and Footers Multi column Long tiny text Base Overall olmOCR pipeline v0.4.0 with olmOCR 2 7B 1025 82.9 82.1 84.3 48.3 95.7 84.3 81.4 99.7 82.3 ± 1.1 olmOCR pipeline v0.4.0 with olmOCR 2 7B 1025 FP8 83.0 82.3 84.9 47.7 96.1 83.7 81.9 99.7 82.4 ± 1.1 Usage This model expects as input a single document image, rendered such that the longest dimension is 1288 pixels. The prompt must th…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy