HUANGCHIHHUNGLeo/claude-real-video
Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, MIT. observed · 2026-08-28
Health v2 · maintenance only
80/100
- Activity 99
- Release rhythm 98
- Longevity 4
Flags: young
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.
- gap_med: 0
- age_days: 64
- days_rel: 15
- days_push: 11
- n_releases_24m: 24
Adoption not part of the score
2072 stars · 180 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
A Python CLI tool and agent skill that lets LLMs like Claude actually watch videos by extracting scene-aware, deduplicated keyframes plus a transcript from a URL or local file. It integrates with Claude Code, Cursor, and other agent hosts via npx skills or a plugin marketplace.
Use cases
- let claude watch a youtube video
- summarize a video with an llm
- extract keyframes from a video for ai analysis
- get a transcript of a video file
- ask questions about a video's visual content
- deduplicate video frames to save tokens
- analyze video content with a multimodal agent
When to choose
- you want an LLM to reason about video visuals, not just the transcript
- you use Claude Code, Cursor, or another agent host and want video understanding as a skill
- you need local, MIT-licensed video frame extraction and transcription
- you want token-efficient frame sampling with contact sheets
When to avoid
- you need a hosted video-understanding API with no local setup
- you only need video transcripts without visual analysis
- you need advanced cinematic analysis like cut rhythm and emotion timelines, which is behind the paid Pro add-on
Facets
cli-tool · maturity active
video-processing cli speech-recognition machine-learning developer-tools artificial-intelligence large-language-models developer-tools python cli cross-platform video-analysis keyframe-extraction scene-detection transcription whisper ffmpeg multimodal claude-code llm-tools agent-skills video command-line
2 sources
- readme: https://github.com/HUANGCHIHHUNGLeo/claude-real-video · fetched 2026-08-28 · 550d9f8479d4
- registry_pypi: https://pypi.org/pypi/claude-real-video/json · fetched 2026-08-29 · 85c98994b2f0
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| HUANGCHIHHUNGLeo/claude-real-video | main | 80 |
For agents
markdown · JSON · MCP: product_card(name="HUANGCHIHHUNGLeo/claude-real-video")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem