DeepSeek
DeepSeek-OCR-2
About
DeepSeek-OCR-2 is a 3.4B open-weight model from DeepSeek. DeepSeek OCR 2 and markdown extraction. It accepts text, images, and documents. Context window is 32K. License is MIT. Released January 27, 2026.
Parameters
3.4B
Published parameter count.
Context window
32K
Tokens of context on a request.
License
MIT
License on the weights.
Pricing
Token rates are US dollars per 1M tokens. Image rates are per 1K images. Audio rates are per 1M audio seconds.
| Rate | Price |
|---|---|
| Input / 1M | $0.25 |
| Cached input / 1M | $0.25 |
| Output / 1M | $0.80 |
Compare
DeepSeek-OCR-2 is one of the catalog's OCR models. The table is each model's published price, size, and context.
| Model | What it does | Price | Parameters | Context |
|---|---|---|---|---|
| DeepSeek-OCR-2 | DeepSeek OCR 2 and markdown extraction. | $0.25 / 1M input | 3.4B | 32K |
| PP-OCRv6 | PaddleOCR PP-OCRv6 medium text detection and recognition; scene OCR JSONL on image chat; text returns plain text; document_url fan-out via text. | $0.01 / 1M input | 20M | — |
| dots.mocr | Multilingual document layout parsing and markdown OCR. | $0.20 / 1M input | 3B | 32K |
| PaddleOCR-VL 1.6 | PaddleOCR-VL-1.6 for OCR, tables, formulas, and charts. | $0.15 / 1M input | 0.9B | 16K |
Benchmarks
Published scores for DeepSeek-OCR-2, from the OmniDocBench leaderboard.
- OmniDocBench v1.690.25
- Formula CDM91.84
- Table TEDS83.89
| Reading | Example |
|---|---|
| Qualitative | Clear structure, grounded in the input |
| Quantitative | OmniDocBench v1.6: 90.25. |
| Cost and performance | Lower listed rate, mid-pack latency |
Methods
Methods this model serves. Payload shapes are in the docs.
| Method | Returns |
|---|---|
| ocr | Lines of text |
| markdown | The page as Markdown |
| free_ocr | Plain text |
| grounding_ocr | Boxes for a phrase you name |
Estimate cost
Estimate. A page is 1,200 input tokens and 1,000 output tokens.
$1.10
Quick start
Call this model on the OpenAI-compatible gateway. The model id is already filled in.
