Ross ROSS = Recommend OSS · open-source software intelligence for agents

zai-org/SCAIL

SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (CVPR 2026 Findings) observed · 2026-08-28

github.com/zai-org/SCAIL · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

52/100

  • Activity 81
  • Release rhythm 35
  • Longevity 19

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 275
  • days_rel: n/a
  • days_push: 119
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1043 stars · 59 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

SCAIL is the official inference implementation of a 14B diffusion transformer model that generates studio-grade character animation videos from a reference character image and a driving pose sequence. It uses in-context learning of 3D-consistent pose representations to handle complex motions like turning and flipping with coherent depth-aware motion.

Use cases

  • animate a character image from a motion capture or pose video
  • generate studio-grade character animation from pose sequences
  • drive a static character with complex motions like flips and turns
  • video-to-video character reenactment with 3D-consistent poses
  • produce production-quality animated character videos without a studio pipeline
  • research controllable video generation with pose conditioning

When to choose

  • you need high-fidelity character animation driven by pose input
  • your motions involve complex body rotation, flipping, or turning that simpler pose-guided models fail on
  • you want to run inference with a pretrained 14B character animation model on GPU
  • you are researching pose representation and conditioning for video diffusion models

When to avoid

  • you lack a high-VRAM GPU to run a 14B diffusion transformer
  • you need training or fine-tuning code rather than inference
  • you need real-time or low-latency animation
  • you want lightweight 2D skeletal animation for games rather than generated video

Facets

library · maturity active

video-processing machine-learning deep-learning llm-inference artificial-intelligence computer-vision deep-learning python character-animation video-generation video2video pose-driven-animation diffusion-transformer cvpr-2026 in-context-learning 3d-pose-representation image-to-video research-model video gpu linux

1 source

Member repositories

RepositoryRoleHealth v2
zai-org/SCAILmain52

For agents

markdown · JSON · MCP: product_card(name="zai-org/SCAIL")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem