Ross ROSS = Recommend OSS · open-source software intelligence for agents

Soul-AILab/SoulX-LiveAct

Official inference code for SoulX-LiveAct: Towards Hour-Scale Real-Time Human Animation with Neighbor Forcing and ConvKV Memory observed · 2026-08-28

github.com/Soul-AILab/SoulX-LiveAct · Python observed · 2026-08-28

Health v2 · maintenance only

54/100

  • Activity 87
  • Release rhythm 35
  • Longevity 12

Flags: no_releases young no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 174
  • days_rel: n/a
  • days_push: 79
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1176 stars · 99 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

SoulX-LiveAct is the official inference code for a real-time human animation framework that generates lifelike, audio/multimodal-controlled talking-head video streams using autoregressive diffusion with Neighbor Forcing and ConvKV memory compression. It achieves 20 FPS on two H100/H200 GPUs and supports consumer GPUs like the RTX 5090 at reduced frame rates.

Use cases

  • generate real-time talking avatar video from audio
  • stream hour-long digital human animations without memory growth
  • run audio-driven human animation on consumer GPUs
  • build a virtual podcast or talk show presenter
  • deploy a real-time digital human for live interaction
  • compress KV cache for long autoregressive video diffusion

When to choose

  • you need real-time, audio-driven talking-head video generation with hour-scale duration
  • you have modern NVIDIA GPUs (H100/H200 or RTX 4090/5090) and want FP8/FP4 optimized inference
  • you want a research-grade AR diffusion video model with constant-memory streaming

When to avoid

  • you need a simple offline video generation pipeline without real-time constraints
  • you lack CUDA GPUs or sufficient VRAM for an 18B-parameter model
  • you need a permissively licensed production dependency - no license is specified

Facets

library · maturity active

llm-inference video-processing machine-learning deep-learning gpu-computing deep-learning artificial-intelligence computer-vision python diffusion-models human-animation talking-head real-time-video-generation autoregressive-video audio-driven-animation digital-human inference-optimization video linux gpu

1 source

Member repositories

RepositoryRoleHealth v2
Soul-AILab/SoulX-LiveActmain54

For agents

markdown · JSON · MCP: product_card(name="Soul-AILab/SoulX-LiveAct")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem