Ross ROSS = Recommend OSS · open-source software intelligence for agents

Soul-AILab/SoulX-FlashTalk

SoulX-FlashTalk is the first 14B model to achieve sub-second start-up latency (0.87s) while maintaining a real-time throughput of 32 FPS on an 8xH800 node. observed · 2026-08-28

github.com/Soul-AILab/SoulX-FlashTalk · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

58/100

  • Activity 95
  • Release rhythm 35
  • Longevity 17

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 251
  • days_rel: n/a
  • days_push: 34
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1479 stars · 138 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

SoulX-FlashTalk is a 14B audio-driven talking avatar model that streams infinite real-time video from a reference image and audio, achieving 0.87s startup latency and 32 FPS on an 8xH800 node. It ships inference code and Hugging Face model weights under Apache-2.0.

Use cases

  • generate a talking avatar video from a photo and audio
  • real-time streaming digital human for live streaming
  • build a video podcast presenter from audio
  • low-latency audio-driven talking head generation
  • infinite-length avatar video generation
  • virtual host for livestreams

When to choose

  • you need real-time, infinite-length audio-driven avatar video with sub-second startup
  • you have multi-GPU (8xH800-class) infrastructure
  • you want an Apache-2.0 model with released weights and inference code

When to avoid

  • you only have a single consumer GPU - use SoulX-FlashHead instead
  • you need training/fine-tuning code, which is not released
  • you need a hosted online demo, which is not yet available

Facets

library · maturity active

video-processing audio-processing machine-learning llm-inference gpu-computing artificial-intelligence computer-vision deep-learning python talking-head audio-driven-avatar streaming-inference video-generation digital-human real-time model-weights inference video audio linux gpu docker

1 source

Member repositories

RepositoryRoleHealth v2
Soul-AILab/SoulX-FlashTalkmain58

For agents

markdown · JSON · MCP: product_card(name="Soul-AILab/SoulX-FlashTalk")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem