π Github π₯ Model Download π Paper Link π Arxiv Paper Link DeepSeek OCR: Contexts Optical Compression Explore the boundaries of visual text compression. Usage Inference using Huggingface transformers on NVIDIA GPUs. Requirements tested on python 3.12.9 + CUDA11.8οΌ vLLM Refer to πGitHub for guidance on model inference acceleration and PDF processing, etc. [2025/10/23] πππ DeepSeek OCR is now officially supported in upstream vLLM. Visualizations Acknowledgement We would like to thank Vary, GOT OCR2.0, MinerU, PaddleOCR, OneChart, Slow Perception for their valuable models and ideas. We also appreciate the benchmarks: Fox, OminiDocBench. Citation bibtex @article{wei2025deepseek, title={DeepSeek OCR: Contexts Optical Compression}, author={Wei, Haoran and Sun, Yaofeng and Li, Yukun}, journal={arXiv preprint arXiv:2510.18234}, year={2025} }
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy