Nanonets OCR2: A model for transforming documents into structured markdown with intelligent content recognition and semantic tagging 🖥️ Live Demo 📢 Blog ⌨️ GitHub 📖 Cookbooks Nanonets OCR2 by Nanonets is a family of powerful, state of the art image to markdown OCR models that go far beyond traditional text extraction. It transforms documents into structured markdown with intelligent content recognition and semantic tagging, making it ideal for downstream processing by Large Language Models (LLMs). Nanonets OCR2 is packed with features designed to handle complex documents with ease: LaTeX Equation Recognition: Automatically converts mathematical equations and formulas into properly formatted LaTeX syntax. It distinguishes between inline ( $...$ ) and display ( $$...$$ ) equations. Intelligent Image Description: Describes images within documents using structured tags, making them digestible for LLM processing. It can describe various image types, including logos, charts, graphs and so on, detailing their content, style, and context. Signature Detection & Isolation: Identifies and isolates signatures from other text, outputting them within a tag. This is crucial for processing lega…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy