Ross ROSS = Recommend OSS · open-source software intelligence for agents

YuanxunLu/LiveSpeechPortraits

Live Speech Portraits: Real-Time Photorealistic Talking-Head Animation (SIGGRAPH Asia 2021) observed · 2026-08-28

github.com/YuanxunLu/LiveSpeechPortraits · Python · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

32/100

  • Activity 0
  • Release rhythm 35
  • Longevity 100

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 1854
  • days_rel: n/a
  • days_push: 1171
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1283 stars · 216 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

A PyTorch implementation of the SIGGRAPH Asia 2021 paper 'Live Speech Portraits', which generates photorealistic personalized talking-head animation from audio input in real time (30+ fps). It combines audio feature extraction, head pose and upper-body motion prediction, and image-to-image translation to synthesize high-fidelity facial renderings.

Use cases

  • generate a talking-head video from an audio clip
  • animate a portrait photo driven by speech audio
  • create real-time photorealistic avatars from voice
  • research on audio-driven facial animation
  • synthesize head poses and facial dynamics from speech
  • reproduce SIGGRAPH Asia 2021 talking-head paper results

When to choose

  • you need audio-driven photorealistic talking-head generation with a specific trained person model
  • you want a research baseline or to build on a published real-time talking-head method
  • you have a GPU environment and can work with research-grade code

When to avoid

  • you need a production-ready product with a polished UI or API
  • you need to animate arbitrary new identities without training data for that person
  • you need actively maintained code with recent dependency support
  • you need full-body or non-portrait avatar animation

Facets

library · maturity maintenance

machine-learning deep-learning video-processing audio-processing computer-vision speech-recognition computer-vision deep-learning artificial-intelligence media python windows talking-head audio-driven-animation photorealistic face-synthesis research-code siggraph video audio linux gpu

1 source

Member repositories

RepositoryRoleHealth v2
YuanxunLu/LiveSpeechPortraitsmain32

For agents

markdown · JSON · MCP: product_card(name="YuanxunLu/LiveSpeechPortraits")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem