Ross ROSS = Recommend OSS · open-source software intelligence for agents

MeiGen-AI/MultiTalk

[NeurIPS 2025] Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation observed · 2026-08-28

github.com/MeiGen-AI/MultiTalk · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

56/100

  • Activity 83
  • Release rhythm 35
  • Longevity 33

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 462
  • days_rel: n/a
  • days_push: 104
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

2992 stars · 497 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

MultiTalk is an audio-driven framework for generating multi-person conversational videos from multi-stream audio, a reference image, and a text prompt, with lip motions synchronized to each speaker. It supports conversation, singing, interaction control, and cartoon-style video generation.

Use cases

  • generate talking head videos from audio
  • create multi-person conversation videos with lip sync
  • make a person sing from an audio track
  • animate a reference photo to speak given dialogue audio
  • generate cartoon character conversation videos
  • control interactions between people in generated video

When to choose

  • you need audio-driven video generation with accurate per-speaker lip sync
  • you want to animate multiple people in one video from separate audio streams
  • you need a research-grade, actively maintained model with published results (NeurIPS 2025)

When to avoid

  • you need real-time or low-latency video generation
  • you lack a GPU or cannot run heavy diffusion models
  • you only need single-image animation without audio alignment

Facets

library · maturity active

video-processing machine-learning deep-learning audio-processing image-processing deep-learning computer-vision artificial-intelligence python talking-head-generation lip-sync audio-driven-video multi-person-conversation video-generation diffusion-models research-code neurips-2025 video audio gpu linux

2 sources

Member repositories

RepositoryRoleHealth v2
MeiGen-AI/MultiTalkmain56

For agents

markdown · JSON · MCP: product_card(name="MeiGen-AI/MultiTalk")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem