Ross ROSS = Recommend OSS · open-source software intelligence for agents

deepseek-ai/DeepSeek-OCR-2

Visual Causal Flow observed · 2026-08-28

github.com/deepseek-ai/DeepSeek-OCR-2 · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

44/100

  • Activity 65
  • Release rhythm 35
  • Longevity 15

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 218
  • days_rel: n/a
  • days_push: 212
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

3379 stars · 304 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

DeepSeek-OCR 2 is an open-source vision-language model and inference toolkit implementing 'Visual Causal Flow' for optical character recognition and document understanding. It provides scripts for running inference on images and PDFs via vLLM or Hugging Face Transformers.

Use cases

  • extract text from images with ocr
  • convert pdf documents to text
  • parse scanned documents
  • run ocr model locally on gpu
  • benchmark document parsing on OmniDocBench
  • batch process images for text extraction

When to choose

  • you need high-quality OCR or document-to-text conversion with a modern vision-language model
  • you want GPU-accelerated batch or concurrent PDF OCR via vLLM
  • you're researching visual encoding approaches like Visual Causal Flow

When to avoid

  • you need lightweight CPU-only OCR without a GPU
  • you need a simple hosted OCR API rather than self-hosted model inference
  • you lack the CUDA/torch environment the project requires

Facets

library · maturity active

ocr machine-learning llm-inference pdf computer-vision image-processing pdf deep-learning artificial-intelligence python vision-language-model document-parsing vllm transformers visual-causal-flow deepseek gpu linux docker

1 source

Member repositories

RepositoryRoleHealth v2
deepseek-ai/DeepSeek-OCR-2main44

For agents

markdown · JSON · MCP: product_card(name="deepseek-ai/DeepSeek-OCR-2")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem