Ross ROSS = Recommend OSS · open-source software intelligence for agents

Saiyan-World/goku

[CVPR2025 Highlight] Video Generation Foundation Models: https://saiyan-world.github.io/goku/ observed · 2026-08-28

github.com/Saiyan-World/goku · homepage · Python observed · 2026-08-28

Health v2 · maintenance only

23/100

  • Activity 7
  • Release rhythm 35
  • Longevity 40

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 570
  • days_rel: n/a
  • days_push: 560
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

2905 stars · 308 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Goku is a family of flow-based (rectified flow Transformer) foundation models for joint image and video generation, released by HKU and ByteDance. The repository provides model code, weights, and inference pipelines for text-to-video, image-to-video, and text-to-image generation.

Use cases

  • generate videos from text prompts
  • animate a still image into a video
  • generate high-quality images from text descriptions
  • research rectified flow transformer models for visual generation
  • benchmark video generation models on VBench or MovieGenBench

When to choose

  • you need state-of-the-art open text-to-video or image-to-video generation
  • you are researching flow-based generative models and want strong baselines
  • you have GPU resources for running large generative models locally

When to avoid

  • you need a lightweight or CPU-only image/video generator
  • you require a permissive license for commercial use (no license is specified)
  • you only need simple video editing rather than generative synthesis

Facets

library · maturity active

machine-learning deep-learning video-processing image-processing llm-inference deep-learning image-processing artificial-intelligence python text-to-video image-to-video text-to-image rectified-flow diffusion-transformer cvpr2025 foundation-model generative-ai video gpu linux

2 sources

Member repositories

RepositoryRoleHealth v2
Saiyan-World/gokumain23

For agents

markdown · JSON · MCP: product_card(name="Saiyan-World/goku")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem