Ross ROSS = Recommend OSS · open-source software intelligence for agents

Tencent-Hunyuan/MixGRPO

[ECCV 2026] MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE observed · 2026-08-28

github.com/Tencent-Hunyuan/MixGRPO · homepage · Python · NOASSERTION (other) observed · 2026-08-28

Health v2 · maintenance only

58/100

  • Activity 90
  • Release rhythm 35
  • Longevity 28

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 400
  • days_rel: n/a
  • days_push: 63
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1177 stars · 51 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

MixGRPO is a research framework from Tencent Hunyuan implementing a mixed ODE-SDE GRPO algorithm for efficient reinforcement learning fine-tuning of flow-based diffusion image generation models. It includes training code, a FLUX.1 Dev fine-tuned checkpoint, and a faster MixGRPO-Flash variant using a sliding window mechanism.

Use cases

  • fine-tune diffusion models with GRPO for human preference alignment
  • speed up flow-based RL post-training of image generators
  • reproduce MixGRPO results from the ECCV 2026 paper
  • train SD3.5 or FLUX LoRA with reward models like HPSv2
  • compare Flow-GRPO and MixGRPO training efficiency

When to choose

  • you need RL-based preference alignment for flow matching image models
  • you want faster GRPO training with fewer denoising steps optimized
  • you want the official implementation and checkpoints of the MixGRPO paper

When to avoid

  • you need a production-ready, well-licensed training framework
  • you are not working with diffusion or flow-based generative models
  • you need stable long-term support rather than research code

Facets

library · maturity active

reinforcement-learning machine-learning llm-training image-processing machine-learning deep-learning image-processing artificial-intelligence python diffusion-models grpo flow-matching rlhf image-generation research-code eccv-2026 gpu linux

2 sources

Member repositories

RepositoryRoleHealth v2
Tencent-Hunyuan/MixGRPOmain58

For agents

markdown · JSON · MCP: product_card(name="Tencent-Hunyuan/MixGRPO")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem