JiuhaiChen/BLIP3o
Official implementation of BLIP3o-Series observed · 2026-08-28
Health v2 · maintenance only
44/100
- Activity 54
- Release rhythm 35
- Longevity 35
Flags: no_releases no_license
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 497
- days_rel: n/a
- days_push: 277
- n_releases_24m: 0
Adoption not part of the score
1666 stars · 79 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
Official implementation of the BLIP3o-Series, a unified autoregressive-plus-diffusion model for text-to-image generation and editing. It combines an autoregressive model that produces intermediate features with a diffusion model for image synthesis, trained with discrete image token supervision and GRPO reinforcement learning.
Use cases
- generate images from text prompts
- train a unified autoregressive diffusion image generation model
- fine-tune a text-to-image model with instruction tuning data
- apply GRPO reinforcement learning to improve prompt alignment in image generation
- improve text rendering in generated images
- reproduce BLIP3o-NEXT research results
- pretrain a multimodal image generation model on open caption datasets
When to choose
- you need an open-source, fully reproducible text-to-image model with training code and data
- you want to experiment with combining autoregressive and diffusion architectures
- you want to apply RLHF-style GRPO training to image generation models
- you are researching discrete image token supervision for multimodal models
When to avoid
- you just want a plug-and-play image generator without training or research work
- you lack GPU/Slurm infrastructure for large model training
- you need a commercially licensed model (no license is specified)
- you need a stable production-ready library rather than research code
Facets
library · maturity active
machine-learning deep-learning llm-training image-processing artificial-intelligence deep-learning image-processing large-language-models python text-to-image autoregressive-diffusion image-generation grpo reinforcement-learning research-code multimodal linux gpu
1 source
- readme: https://github.com/JiuhaiChen/BLIP3o · fetched 2026-08-28 · 249f703be707
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| JiuhaiChen/BLIP3o | main | 44 |
For agents
markdown · JSON · MCP: product_card(name="JiuhaiChen/BLIP3o")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem