Dedicated model · Available as managed deployment
DeepSeek's 3B document model — OCR that turns scanned pages, PDFs and images into text and markdown with tables and layout, under MIT. Validated on AxForge hardware and deployed on a dedicated DGX Spark for your traffic only — an OpenAI-compatible endpoint on hardware only you use, operated by AxForge in the EU.
Why AxForge
| Documents in, text out | Scanned pages, forms, invoices and books become clean text and markdown — with tables, headings and reading order preserved. |
|---|---|
| Thousands of pages an hour | A 3B model on a dedicated DGX Spark processes documents at a rate a general vision model cannot match, at a fraction of the cost. |
| Stays in the EU | Contracts, medical records and customer documents never leave your dedicated machine — no third-party OCR API, zero retention. |
Specifications
| Model | DeepSeek-OCR — deepseek-ai |
|---|---|
| Modalities | Text + vision — accepts image input |
| Sizes | 3.3B |
| Context window | 8,192 tokens |
| Licence | Open weights — mit; commercial use permitted |
| Hardware | NVIDIA DGX Spark (GB10, 128 GB unified memory) — owned and operated by AxForge |
| Rental term | Hour, week, month or year |
| Hardware pricing | €0.69/hour on demand · €0.66/hour by the week · €0.62/hour by the month · €0.55/hour by the year, excl. VAT |
| Managed service | Quoted per deployment |
| Region | Málaga, Spain (eu-es-1) |
Full details, benchmarks and FAQ on the DeepSeek-OCR page. Prices exclude VAT.
How it works
| 1 | Request deployment — describe your traffic, context needs and rental term. |
|---|---|
| 2 | You receive the configuration, hardware rental and managed-service price in writing before anything is billed. |
| 3 | AxForge deploys DeepSeek-OCR on a dedicated DGX Spark reserved for you. |
| 4 | Point your OpenAI SDK at your own endpoint with the model name you receive. |
| 5 | Adjust the term — hour, week, month or year — as your workload settles. |
Request deployment or sign in to start.
FAQ
Not on the serverless API — it is available as a managed deployment: validated on AxForge hardware and deployed on a dedicated DGX Spark for your traffic only. The serverless API serves Qwen3.8 27B.
It is built for one job — reading documents into text — and does it faster and cheaper than a general vision model; for questions about the content, pair it with an LLM on the same machine.
Images and page renders of PDFs; AxForge sets up the page-to-image step with the deployment so you send documents and receive markdown.
AxForge publishes only numbers it measures itself, and has not benchmarked this model on its nodes yet. For quality benchmarks, see the official model card.
Hardware by the hour, week, month or year; the managed service is quoted per deployment — both confirmed in writing before anything is billed.