# EBazarov/nsfw_data_source_urls

Collection of NSFW images URLs for the purposes of training an NSFW Image Classifier

Repository: https://github.com/EBazarov/nsfw_data_source_urls
Canonical: https://ross.abutalabs.com/products/nsfw_data_source_urls
License: MIT
License Family: permissive
Last push: 2020-12-14T09:40:00+00:00

## Health v2 (maintenance only)
Score: 32/100 (v2, computed 2026-09-02T17:46:02.011165+00:00)
- activity 0, release rhythm 35, longevity 100
- inputs: {"age_days": 2758, "days_push": 2088, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 3577, forks 747 (observed 2026-08-28T04:08:10.826675+00:00)

## What it is
A curated collection of text files listing over 1.5 million image URLs across 159 categories, intended for downloading and building NSFW image classification datasets. It provides raw URL lists rather than the images themselves, with suggested use of external scraping scripts.

## Use cases
- train an NSFW image classifier
- build a content moderation dataset
- download labeled adult image data for model training
- create a safe-for-work vs not-safe-for-work training set
- gather image URLs for computer vision research

## When to choose
- you need large-scale labeled image URLs to train an NSFW detection model
- you want category-organized raw data for building a custom image classifier

## When to avoid
- you need the actual images rather than URLs
- you need a maintained, regularly updated dataset
- you need a ready-to-use trained model

## Facets
- artifact type: dataset
- maturity: maintenance
- function: machine-learning, data-generation, web-scraping
- domain: machine-learning, computer-vision, artificial-intelligence
- platform: cross-platform
- tags: nsfw-detection, image-classification, image-urls, training-data, content-moderation

## Member repositories
- EBazarov/nsfw_data_source_urls (main) score 32

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:08:10.826675+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-29T18:34:02.155812+00:00, confidence not recorded.
  - readme: https://github.com/EBazarov/nsfw_data_source_urls (fetched 2026-08-28T04:08:10.826675+00:00, sha 0c94014139df)
- Data as of 2026-08-30T08:39:29.467469+00:00.
