chandra ocr 2 GGUF Chandra OCR 2 from Datalab is a state of the art OCR model that outputs structured markdown, HTML, or JSON while preserving precise layout information from images and PDFs across 90+ languages. It achieves SOTA benchmarks with 85.9% on olmocr and 77.8% multilingual score (+12% over Chandra 1), delivering major gains in math equation parsing, complex table reconstruction (including merged cells), handwriting recognition, form elements like checkboxes, and wide document layouts alongside vastly improved image captioning and diagram extraction. Available via free playground, hosted API for production speed/accuracy, or local deployment through HuggingFace Transformers/vLLM, it excels at transforming challenging real world documents—financial filings, research papers, historical scans, multilingual forms—into semantically rich structured data for downstream AI pipelines and automation workflows. Model Files File Name Quant Type File Size File Link chandra ocr 2.BF16.gguf BF16 9.7 GB Download chandra ocr 2.F16.gguf F16 9.7 GB Download chandra ocr 2.Q2 K.gguf Q2 K 2.12 GB Download chandra ocr 2.Q3 K L.gguf Q3 K L 2.69 GB Download chandra ocr 2.Q3 K M.gguf Q3 K M 2.54 G…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy