bellingcat/auto-archiver
Automatically archive links to videos, images, and social media content from Google Sheets (and more). observed · 2026-08-28
Health v2 · maintenance only
92/100
- Activity 98
- Release rhythm 81
- Longevity 100
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: 5
- age_days: 2056
- days_rel: 128
- days_push: 15
- n_releases_24m: 20
Adoption not part of the score
1109 stars · 107 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A Python tool by Bellingcat that automatically archives web content such as videos, images, social media posts, and webpages from URLs supplied via Google Sheets, CSV files, or the command line. Archived content can be enriched and stored locally or remotely (S3, Google Drive), with status reports written back to the source.
Use cases
- archive social media posts before they get deleted
- bulk archive URLs from a Google Sheet
- preserve videos and images for open-source investigations
- automatically snapshot webpages for evidence preservation
- download and store online content for OSINT research
When to choose
- you need verifiable, automated archiving of online content for research or journalism
- you want to feed URLs from Google Sheets or CSV and get status back
- you want a Docker-deployable archiving pipeline with remote storage options
When to avoid
- you need a browser-based one-click webpage archiver like the Wayback Machine
- you only need simple link bookmarking without content capture
- your sources require heavy JavaScript interaction beyond the supported extractors
Facets
cli-tool · maturity active
web-scraping cli etl file-upload osint crawlers developer-tools python cli cross-platform web-archiving social-media-archiving google-sheets osint bellingcat evidence-preservation automation docker
3 sources
- readme: https://github.com/bellingcat/auto-archiver · fetched 2026-08-28 · a9dc3bddcb82
- homepage: https://pypi.org/project/auto-archiver/ · fetched 2026-08-29 · 4b4e8fead74a
- registry_pypi: https://pypi.org/pypi/auto-archiver/json · fetched 2026-08-29 · 1cc9da65614b
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| bellingcat/auto-archiver | main | 92 |
For agents
markdown · JSON · MCP: product_card(name="bellingcat/auto-archiver")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem