# KaiDMML/FakeNewsNet

This is a dataset for fake news detection research

Repository: https://github.com/KaiDMML/FakeNewsNet
Canonical: https://ross.abutalabs.com/products/fakenewsnet
Language: Python
License Family: other
Last push: 2022-12-08T04:54:39+00:00

## Health v2 (maintenance only)
Score: 32/100 (v2, computed 2026-09-03T02:20:16.233290+00:00)
- activity 0, release rhythm 35, longevity 100
- inputs: {"age_days": 3317, "days_push": 1364, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases, no_license
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1346, forks 475 (observed 2026-08-28T04:04:27.429894+00:00)

## What it is
FakeNewsNet is a research dataset for fake news detection, with minimal CSV files of news samples from PolitiFact and GossipCop plus Python scripts to download full article text and Twitter engagement data. It is intended for academic research on misinformation detection.

## Use cases
- download a fake news detection dataset for research
- collect news articles and tweet data for misinformation studies
- train machine learning models to classify fake vs real news
- benchmark fake news detection algorithms
- get labeled news samples from PolitiFact and GossipCop
- study social media sharing of fake news

## When to choose
- you need a labeled fake vs real news benchmark for ML research
- you want news article text plus Twitter engagement data
- you are studying misinformation propagation on social media

## When to avoid
- you need the complete raw dataset without running collection scripts yourself
- you cannot obtain Twitter API keys
- you need a production-ready misinformation detection service rather than a dataset

## Facets
- artifact type: dataset
- maturity: maintenance
- function: data-generation, web-scraping, nlp
- domain: social-media, data-science, machine-learning
- platform: python, cross-platform
- tags: fake-news-detection, misinformation, twitter-data, benchmark-dataset, news-media, research-dataset, natural-language-processing

## Member repositories
- KaiDMML/FakeNewsNet (main) score 32

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:04:27.429894+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T04:42:40.354071+00:00, confidence not recorded.
  - readme: https://github.com/KaiDMML/FakeNewsNet (fetched 2026-08-28T04:04:27.429894+00:00, sha e247620fa506)
- Data as of 2026-08-30T08:39:29.467469+00:00.
