philipperemy/name-dataset resource
The Python library for names. observed · 2026-08-28
Health v2 · maintenance only
39/100
- Activity 15
- Release rhythm 35
- Longevity 100
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 2885
- days_rel: n/a
- days_push: 512
- n_releases_24m: 0
Adoption not part of the score
1017 stars · 155 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A Python library and dataset of 730K first names and 983K last names with country rankings, gender prediction, fuzzy search, and autocomplete. Data was extracted from a large Facebook user dump and covers 105 countries.
Use cases
- predict gender from a first name
- guess the country of origin of a name
- validate and enrich user names in a signup form
- fuzzy match misspelled names
- autocomplete name input fields
- analyze name popularity by country
When to choose
- you need offline name-to-gender or name-to-country inference in Python
- you want a large curated name dataset for analysis or ML features
- you need fuzzy or prefix search over first and last names
When to avoid
- your environment cannot spare ~3.2 GB of RAM
- you need real-time lookups with fast startup
- you require verified or consented data sources
Facets
dataset · maturity active
nlp data-science search-engine data-science analytics python cross-platform names gender-prediction name-classification fuzzy-search autocomplete demographics natural-language-processing
1 source
- readme: https://github.com/philipperemy/name-dataset · fetched 2026-08-28 · 5ebe3aeb5b4c
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| philipperemy/name-dataset | main | 39 |
For agents
markdown · JSON · MCP: product_card(name="philipperemy/name-dataset")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem