๐ข New Release: Weโve released granite docling 258M , the successor to SmolDocling . It will now receive updates and support, check it out! SmolDocling 256M preview SmolDocling is a multimodal Image Text to Text model designed for efficient document conversion. It retains Docling's most popular features while ensuring full compatibility with Docling through seamless support for DoclingDocuments . This model was presented in the paper SmolDocling: An ultra compact vision language model for end to end multi modal document conversion. ๐ Features: ๐ท๏ธ DocTags for Efficient Tokenization โ Introduces DocTags an efficient and minimal representation for documents that is fully compatible with DoclingDocuments . ๐ OCR (Optical Character Recognition) โ Extracts text accurately from images. ๐ Layout and Localization โ Preserves document structure and document element bounding boxes . ๐ป Code Recognition โ Detects and formats code blocks including identation. ๐ข Formula Recognition โ Identifies and processes mathematical expressions. ๐ Chart Recognition โ Extracts and interprets chart data. ๐ Table Recognition โ Supports column and row headers for structured table extraction. ๐ผ๏ธ Figure Cโฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy