# wanghaisheng/awesome-ocr

A curated list of promising OCR resources

Repository: https://github.com/wanghaisheng/awesome-ocr
Canonical: https://ross.abutalabs.com/products/wanghaisheng-awesome-ocr
Homepage: http://wanghaisheng.github.io/ocr-arxiv-daily/
License: MIT
License Family: permissive
Last push: 2022-06-24T04:50:59+00:00

## Health v2 (maintenance only)
Score: 32/100 (v2, computed 2026-09-02T17:46:02.011165+00:00)
- activity 0, release rhythm 35, longevity 100
- inputs: {"age_days": 3793, "days_push": 1531, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1703, forks 351 (observed 2026-08-28T04:05:24.583740+00:00)

## What it is
A curated list of OCR (optical character recognition) resources, including libraries, cloud APIs, tools, and research paper trackers. It also links to companion projects like an arXiv daily paper tracker and tweet monitoring for OCR research.

## Use cases
- find OCR libraries and tools for a project
- compare cloud OCR APIs like Baidu, Tencent, and Aliyun
- track the latest OCR research papers on arXiv
- find open-source text recognition SDKs like Tesseract or PaddleOCR
- discover resources for recognizing Chinese text from images
- learn about OCR engines for historical document processing

## When to choose
- you want a starting point to survey the OCR ecosystem
- you need to discover libraries, APIs, and research in one place
- you want to follow new OCR papers automatically

## When to avoid
- you need a production-ready OCR engine itself rather than a resource list
- you need actively maintained code - this is a curated list, not software
- you need up-to-date API pricing details, which may be stale

## Facets
- artifact type: learning-resource
- maturity: maintenance
- function: ocr, nlp, image-processing, developer-tools
- domain: computer-vision, awesome-lists, tutorials
- platform: cross-platform
- tags: awesome-list, ocr-resources, curated-list, research-papers, text-recognition, natural-language-processing

## Member repositories
- wanghaisheng/awesome-ocr (main) score 32

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:05:24.583740+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T03:37:38.872491+00:00, confidence not recorded.
  - readme: https://github.com/wanghaisheng/awesome-ocr (fetched 2026-08-28T04:05:24.583740+00:00, sha 0a74416f16be)
  - homepage: http://wanghaisheng.github.io/ocr-arxiv-daily/ (fetched 2026-08-29T11:12:30.216285+00:00, sha 46bcaabbc1ee)
- Data as of 2026-08-30T08:39:29.467469+00:00.
