Ross ROSS = Recommend OSS · open-source software intelligence for agents

HUANGCHIHHUNGLeo/claude-real-video

Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, MIT. observed · 2026-08-28

github.com/HUANGCHIHHUNGLeo/claude-real-video · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

80/100

  • Activity 99
  • Release rhythm 98
  • Longevity 4

Flags: young

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: 0
  • age_days: 64
  • days_rel: 15
  • days_push: 11
  • n_releases_24m: 24

Full methodology

Adoption not part of the score

2072 stars · 180 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A Python CLI tool and agent skill that lets LLMs like Claude actually watch videos by extracting scene-aware, deduplicated keyframes plus a transcript from a URL or local file. It integrates with Claude Code, Cursor, and other agent hosts via npx skills or a plugin marketplace.

Use cases

  • let claude watch a youtube video
  • summarize a video with an llm
  • extract keyframes from a video for ai analysis
  • get a transcript of a video file
  • ask questions about a video's visual content
  • deduplicate video frames to save tokens
  • analyze video content with a multimodal agent

When to choose

  • you want an LLM to reason about video visuals, not just the transcript
  • you use Claude Code, Cursor, or another agent host and want video understanding as a skill
  • you need local, MIT-licensed video frame extraction and transcription
  • you want token-efficient frame sampling with contact sheets

When to avoid

  • you need a hosted video-understanding API with no local setup
  • you only need video transcripts without visual analysis
  • you need advanced cinematic analysis like cut rhythm and emotion timelines, which is behind the paid Pro add-on

Facets

cli-tool · maturity active

video-processing cli speech-recognition machine-learning developer-tools artificial-intelligence large-language-models developer-tools python cli cross-platform video-analysis keyframe-extraction scene-detection transcription whisper ffmpeg multimodal claude-code llm-tools agent-skills video command-line

2 sources

Member repositories

RepositoryRoleHealth v2
HUANGCHIHHUNGLeo/claude-real-videomain80

For agents

markdown · JSON · MCP: product_card(name="HUANGCHIHHUNGLeo/claude-real-video")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem