Typhoon OCR 3B : A bilingual document parsing model built specifically for real world documents in Thai and English inspired by models like olmOCR based on Qwen2.5 VL Instruction. By using this model, you agree to the OpenTyphoon Terms and Conditions and acknowledge the Privacy Notice: https://opentyphoon.ai/tac · https://opentyphoon.ai/privacy Try our demo available on Demo Code / Examples available on Github Release Blog available on OpenTyphoon Blog Technical Report on Arxiv Remark: This model is intended to be used with a specific prompt only; it will not work with any other prompts. Remark: If you want to run the model locally, we recommend using the Ollama build at https://ollama.com/scb10x. We’ve found that the GGUF files for llama.cpp or LM Studio may suffer from accuracy issues. Real World Document Support 1. Structured Documents : Financial reports, Academic papers, Books, Government forms Output format : Markdown for general text HTML for tables (including merged cells and complex layouts) Figures, charts, and diagrams are represented using figure tags for structured visual understanding Each figure undergoes multi layered interpretation : Observation : Detects elements…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy