baidu/DuReader resource
Baseline Systems of DuReader Dataset observed · 2026-08-28
Health v2 · maintenance only
32/100
- Activity 0
- Release rhythm 35
- Longevity 100
Flags: no_releases no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 3219
- days_rel: n/a
- days_push: 1560
- n_releases_24m: 0
Adoption not part of the score
1178 stars · 306 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
DuReader is a collection of Chinese machine reading comprehension and question answering benchmark datasets, including MRC, passage retrieval, DocVQA, robustness, and question matching variants. The repository also provides baseline systems and models such as KT-NET and D-NET for these benchmarks.
Use cases
- evaluate machine reading comprehension models on Chinese question answering
- benchmark passage retrieval systems on a large-scale Chinese dataset
- test robustness of question matching models against linguistic perturbations
- train and evaluate open-domain document visual question answering models
- compare MRC model generalization across datasets
- download Chinese QA benchmark datasets for research
When to choose
- you need Chinese-language QA or MRC benchmark datasets with leaderboards
- you want baseline models for reading comprehension research
- you are evaluating model robustness, generalization, or opinion polarity judgment in QA
When to avoid
- you need English-language QA benchmarks only
- you want a production-ready QA system rather than research datasets and baselines
- you need actively maintained code with a clear license
Facets
dataset · maturity maintenance
machine-learning nlp search-engine benchmarking machine-learning artificial-intelligence python question-answering machine-reading-comprehension chinese-nlp docvqa passage-retrieval baseline-models natural-language-processing search
2 sources
- readme: https://github.com/baidu/DuReader · fetched 2026-08-28 · a32bb21627fa
- homepage: http://ai.baidu.com/broad/subordinate?dataset=dureader · fetched 2026-08-29 · 379c350ffcca
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| baidu/DuReader | main | 32 |
For agents
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem