# kangvcar/InfoSpider

INFO-SPIDER 是一个集众多数据源于一身的爬虫工具箱🧰，旨在安全快捷的帮助用户拿回自己的数据，工具代码开源，流程透明。支持数据源包括GitHub、QQ邮箱、网易邮箱、阿里邮箱、新浪邮箱、Hotmail邮箱、Outlook邮箱、京东、淘宝、支付宝、中国移动、中国联通、中国电信、知乎、哔哩哔哩、网易云音乐、QQ好友、QQ群、生成朋友圈相册、浏览器浏览历史、12306、博客园、CSDN博客、开源中国博客、简书。

Repository: https://github.com/kangvcar/InfoSpider
Canonical: https://ross.abutalabs.com/products/infospider
Homepage: https://infospider.vercel.app
Language: Python
License: GPL-3.0
License Family: copyleft
Topics: python3, crawl, spider, selenium, wxpython, tkinter, automation, hotmail, chrome, csdn, outlook
Last push: 2026-04-21T21:17:06+00:00

## Health v2 (maintenance only)
Score: 58/100 (v2, computed 2026-09-03T02:20:16.233290+00:00)
- activity 78, release rhythm 8, longevity 100
- inputs: {"age_days": 2244, "days_push": 134, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: none
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 8247, forks 1477 (observed 2026-08-28T04:10:19.842255+00:00)

## What it is
InfoSpider is an open-source Python toolbox that crawls a user's own personal data from dozens of Chinese and international services (email providers, e-commerce, social media, blogs, browsers) using Selenium automation. It aggregates, analyzes, and visualizes the exported data so users can reclaim and understand their digital footprint.

## Use cases
- download all my personal data from taobao and jd
- export my qq mail and outlook emails locally
- scrape my own zhihu and bilibili history
- backup my csdn and jianshu blog posts
- visualize my browsing history and online activity
- reclaim personal data from chinese web services
- aggregate my data spread across many websites

## When to choose
- you want a transparent, open-source way to export your own data from many supported Chinese platforms
- you want a GUI tool that automates login and scraping via Selenium
- you want your personal data aggregated and visualized in one place

## When to avoid
- you need to scrape data belonging to other users or at scale
- you need a headless server-side scraping framework rather than a desktop GUI app
- a target site is not in the supported list and you need custom scraping logic

## Facets
- artifact type: application
- maturity: active
- function: web-scraping, data-visualization, workflow-automation, gui
- domain: privacy, crawlers, data-science
- platform: windows, python, cross-platform
- tags: personal-data, selenium, data-portability, scraper-toolbox, tkinter, wxpython, automation, desktop

## Member repositories
- kangvcar/InfoSpider (main) score 58

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:10:19.842255+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-29T17:29:28.478539+00:00, confidence not recorded.
  - readme: https://github.com/kangvcar/InfoSpider (fetched 2026-08-28T04:10:19.842255+00:00, sha cb16393ad200)
  - homepage: https://infospider.vercel.app (fetched 2026-08-29T08:27:55.920195+00:00, sha 1cd10c6d44e2)
- Data as of 2026-08-30T08:39:29.467469+00:00.
