deepseek-ai/DeepSeek-OCR
Contexts Optical Compression observed · 2026-08-28
Health v2 · maintenance only
45/100
- Activity 64
- Release rhythm 35
- Longevity 22
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 320
- days_rel: n/a
- days_push: 218
- n_releases_24m: 0
Adoption not part of the score
23855 stars · 2203 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded
DeepSeek-OCR is an open vision-language model from DeepSeek AI that researches 'contexts optical compression' - encoding long text contexts as images through a vision encoder feeding an LLM decoder. The repository ships model weights, inference code for both Transformers and vLLM, and an accompanying arXiv paper.
Use cases
- extract text from scanned documents or images with an open-source OCR model
- compress long text contexts into images to extend LLM effective context
- run document understanding with a vision-language model
- serve an OCR model at scale with vLLM
- experiment with optical context compression research from the paper
- parse layouts, tables, and text from PDFs or screenshots
When to choose
- you need high-accuracy open-weights OCR or document parsing with GPU acceleration
- you want to reproduce or build on the contexts-optical-compression research
- you want a DeepSeek vision model with official upstream vLLM support
When to avoid
- you need a lightweight CPU-only OCR utility - this requires CUDA GPUs, flash-attention, and a heavy PyTorch/vLLM stack
- you want the newest iteration - DeepSeek-OCR2 was announced in January 2026 as the successor
- you need a turnkey OCR SaaS with preprocessing, batching UI, and export pipelines
Facets
library · maturity active
ocr machine-learning llm-inference image-processing computer-vision transformers artificial-intelligence machine-learning deep-learning large-language-models computer-vision image-processing python vision-language-model vision-encoder optical-context-compression context-compression model-weights inference-code vllm huggingface multimodal research-model deepseek mit-license natural-language-processing gpu linux
1 source
- readme: https://github.com/deepseek-ai/DeepSeek-OCR · fetched 2026-08-28 · 4ce77cdea615
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| deepseek-ai/DeepSeek-OCR | main | 45 |
For agents
markdown · JSON · MCP: product_card(name="deepseek-ai/DeepSeek-OCR")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem