Ross ROSS = Recommend OSS · open-source software intelligence for agents

HumanAIGC/EMO

Emote Portrait Alive: Generating Expressive Portrait Videos with Audio2Video Diffusion Model under Weak Conditions observed · 2026-08-28

github.com/HumanAIGC/EMO observed · 2026-08-28

Health v2 · maintenance only

25/100

  • Activity 0
  • Release rhythm 35
  • Longevity 65

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 918
  • days_rel: n/a
  • days_push: 742
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

7594 stars · 928 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

EMO (Emote Portrait Alive) is a research codebase from Alibaba's Institute for Intelligent Computing that generates expressive talking portrait videos from a single reference image and an audio clip, using an audio2video diffusion model under weak conditions. It was published at ECCV 2024 and is primarily a paper companion release.

Use cases

  • generate a talking head video from a photo and audio
  • animate a portrait image to match speech audio
  • create expressive avatar videos with audio-driven diffusion
  • reproduce ECCV 2024 audio2video portrait research
  • build lip-synced character videos from voice recordings

When to choose

  • you need state-of-the-art audio-driven portrait video generation for research
  • you want to study or extend the EMO diffusion approach
  • you have GPU resources and are comfortable with research-grade code

When to avoid

  • you need a production-ready tool with a stable API or license
  • you lack a GPU or cannot run heavy diffusion inference
  • you need commercial usage rights - the repo has no license
  • you want a polished end-user application rather than research code

Facets

library · maturity experimental

machine-learning deep-learning video-processing audio-processing computer-vision speech-recognition artificial-intelligence deep-learning computer-vision python audio2video diffusion-model talking-head portrait-animation research-code eccv-2024 no-license video audio gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
HumanAIGC/EMOmain25

For agents

markdown · JSON · MCP: product_card(name="HumanAIGC/EMO")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem