Baidu
Unlimited-OCR
About
Unlimited-OCR is a 3B open-weight model from Baidu. Long-horizon document parsing with batched multi-page OCR. It accepts text, images, and documents. Context window is 32K. License is MIT. Released June 19, 2026.
Parameters
3B
Published parameter count.
Context window
32K
Tokens of context on a request.
License
MIT
License on the weights.
Pricing
Token rates are US dollars per 1M tokens. Image rates are per 1K images. Audio rates are per 1M audio seconds.
| Rate | Price |
|---|---|
| Input / 1M | $0.25 |
| Cached input / 1M | $0.03 |
| Output / 1M | $0.55 |
Compare
Unlimited-OCR is one of the catalog's OCR models. The table is each model's published price, size, and context.
| Model | What it does | Price | Parameters | Context |
|---|---|---|---|---|
| Unlimited-OCR | Long-horizon document parsing with batched multi-page OCR. | $0.25 / 1M input | 3B | 32K |
| PP-OCRv6 | PaddleOCR PP-OCRv6 medium text detection and recognition; scene OCR JSONL on image chat; text returns plain text; document_url fan-out via text. | $0.01 / 1M input | 20M | — |
| DeepSeek-OCR-2 | DeepSeek OCR 2 and markdown extraction. | $0.25 / 1M input | 3.4B | 32K |
| dots.mocr | Multilingual document layout parsing and markdown OCR. | $0.20 / 1M input | 3B | 32K |
Benchmarks
Published scores for Unlimited-OCR, from the Unlimited-OCR paper.
- OmniDocBench v1.693.92
- Formula CDM95.79
- Table TEDS90.16
| Reading | Example |
|---|---|
| Qualitative | Clear structure, grounded in the input |
| Quantitative | OmniDocBench v1.6: 93.92. |
| Cost and performance | Lower listed rate, mid-pack latency |
Methods
Methods this model serves. Payload shapes are in the docs.
| Method | Returns |
|---|---|
| markdown | The page as Markdown |
| multi_page | Markdown across a window of pages |
Estimate cost
Estimate. A page is 2,500 input tokens and 1,000 output tokens.
$1.18
Quick start
Call this model on the OpenAI-compatible gateway. The model id is already filled in.
