Alpha-VLLM/Lumina-DiMOO
Lumina-DiMOO - An Open-Sourced Multi-Modal Large Diffusion Language Model observed · 2026-08-28
Health v2 · maintenance only
55/100
- Activity 83
- Release rhythm 35
- Longevity 25
Flags: no_releases
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: n/a
- age_days: 358
- days_rel: n/a
- days_push: 107
- n_releases_24m: 0
Adoption not part of the score
1015 stars · 63 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outputs. It supports text-to-image generation, image-to-image tasks like editing and inpainting, and image understanding, with training and inference code and checkpoints released.
Use cases
- generate images from text prompts with a diffusion language model
- edit images or do subject-driven generation and inpainting
- understand and answer questions about images with a multimodal model
- research discrete diffusion models for multimodal learning
- speed up multimodal sampling with caching and test-time scaling
- fine-tune a unified multimodal model with GRPO-style training
When to choose
- you need a unified model for both multimodal understanding and generation
- you want faster sampling than autoregressive or hybrid AR-diffusion models
- you are researching discrete diffusion language models
- you want an Apache-2.0 licensed open multimodal foundation model
When to avoid
- you need a lightweight model for CPU-only or edge deployment
- you only need a production chatbot with mature ecosystem tooling
- you require long-form text-only generation with standard AR LLMs
Facets
library · maturity active
machine-learning deep-learning image-processing llm-inference llm-training artificial-intelligence large-language-models image-processing computer-vision deep-learning python discrete-diffusion multimodal text-to-image image-editing diffusion-language-model image-understanding gpu linux
2 sources
- readme: https://github.com/Alpha-VLLM/Lumina-DiMOO · fetched 2026-08-28 · 7b3cc726733c
- homepage: https://synbol.github.io/Lumina-DiMOO/ · fetched 2026-08-29 · f740289fa2be
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| Alpha-VLLM/Lumina-DiMOO | main | 55 |
For agents
markdown · JSON · MCP: product_card(name="Alpha-VLLM/Lumina-DiMOO")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem