Skip to main content

Overview

DeepSeek-OCR is designed for document recognition and image-to-text scenarios, pushing the limits of visual and textual compression. The model can render long texts into highly compressed images, achieving an OCR accuracy of 97% at a lossless compression ratio of 10x and around 60% accuracy at 20x compression.
This model currently supports only single-turn, independent recognition tasks. It does not support multi-turn conversations. Only one image can be uploaded per request, and it is strongly recommended to use preset prompts for optimal performance.

Usage Example

This example uses the <|grounding|>OCR this image. preset prompt to perform image text recognition.
Example input image:
OCR Example Image
Example output:
Last modified on October 25, 2025