function: stable-diffusion
161 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| ControlNet ControlNet is a neural network architecture that adds conditional control (edges, poses, depth, etc.) to pretrained text-to-image diffusion… | 31 | 34091 | maintenance |
| ddean2009/MoneyPrinterPlus A Python desktop application that uses AI LLMs to batch-generate short videos with one click, including automatic video mashup/remixing and… | 30 | 7008 | active |
| leejet/stable-diffusion.cpp A pure C/C++ inference engine for diffusion models (Stable Diffusion, FLUX, Wan, Qwen Image, Z-Image, and more) built on ggml in the style … | 91 | 6846 | active |
| Doubiiu/ToonCrafter ToonCrafter is a generative model that interpolates two cartoon images into a short animation by leveraging pre-trained image-to-video diff… | 29 | 6003 | stable |
| aidlearning/AidLearning-FrameWork AidLux (originally AidLearning) is an AIoT development platform that runs a native Ubuntu Linux environment with GUI, deep learning tooling… | 70 | 5797 | active |
| xxlong0/Wonder3D Wonder3D is a cross-domain diffusion model that reconstructs high-fidelity textured 3D meshes from a single image in 2-3 minutes. It genera… | 32 | 5425 | active |
| yisol/IDM-VTON Official implementation of IDM-VTON, an ECCV 2024 paper that improves diffusion models for high-fidelity virtual try-on, swapping garments … | 30 | 5156 | active |
| tyxsspa/AnyText AnyText is the official implementation of a diffusion-based model for multilingual visual text generation and editing in images, accepted a… | 32 | 4874 | active |
| xlite-dev/lite.ai.toolkit A lightweight C++ toolkit providing unified APIs for 100+ pre-trained AI models across inference backends like ONNX Runtime, MNN, TensorRT,… | 74 | 4427 | active |
| dreamgaussian/dreamgaussian DreamGaussian is the official PyTorch implementation of an ICLR 2024 Oral paper for efficient 3D content creation using generative Gaussian… | 18 | 4352 | active |
| IDEA-CCNL/Fengshenbang-LM Fengshenbang-LM is an open-source suite of Chinese large language models and a PyTorch training framework from IDEA Research's CCNL lab, ai… | 71 | 4123 | active |
| Kosinkadink/ComfyUI-AnimateDiff-Evolved A ComfyUI custom node pack providing an improved AnimateDiff integration plus advanced 'Evolved Sampling' options for animated video genera… | 70 | 3531 | active |
| MooreThreads/Moore-AnimateAnyone An open-source reproduction of AnimateAnyone that animates a character from a single reference image using pose sequences from a driving vi… | 26 | 3514 | active |
| CompVis/latent-diffusion The official research code and pretrained model zoo for Latent Diffusion Models (LDM), the paper behind Stable Diffusion, enabling high-res… | 32 | 14133 | maintenance |
| dirk1983/deepseek A lightweight PHP demo application that calls DeepSeek/OpenAI-compatible chat APIs with streaming (EventSource) output, supporting Markdown… | 36 | 3103 | active |
| sonos/tract Tract is Sonos' tiny, self-contained neural-network inference engine written in Rust. It loads ONNX, TensorFlow/TFLite, and NNEF models, op… | 99 | 3045 | active |
| off-grid-ai/OGAM Off Grid AI (OGAM) is a cross-platform mobile and desktop application that runs AI entirely on-device: GGUF LLM chat with vision, Whisper s… | 78 | 3002 | active |
| TMElyralab/MuseV MuseV is a diffusion-based framework for generating high-fidelity virtual human videos of infinite length using a Visual Conditioned Parall… | 25 | 2846 | active |
| bytedance/InfiniteYou InfiniteYou (InfU) is a research framework from ByteDance for identity-preserved text-to-image generation built on Diffusion Transformers l… | 37 | 2685 | active |
| IceClear/StableSR StableSR is a Python research library that leverages pre-trained Stable Diffusion priors for real-world blind image super-resolution. It pr… | 21 | 2668 | stable |
| koishijs/novelai-bot A Koishi chatbot plugin that generates images via NovelAI, with support for SD-WebUI and Stable Horde backends. It offers model/sampler/siz… | 36 | 2550 | active |
| yifan123/flow_grpo Flow-GRPO is the official PyTorch implementation of a NeurIPS 2025 paper that trains flow matching models (e.g., SD3.5, FLUX.1, Qwen-Image,… | 56 | 2498 | active |
| learnhouse/learnhouse LearnHouse is a next-generation open-source learning management system (LMS) for creating, sharing, and selling educational content. It com… | 99 | 2203 | active |
| SUDO-AI-3D/zero123plus Zero123++ is a diffusion base model that generates consistent multi-view images from a single input image, intended as a stepping stone for… | 27 | 2095 | active |
| storytold/artcraft ArtCraft is an open-source desktop application for interactive AI image and video creation, described as 'the IDE for artists'. It provides… | 94 | 2044 | active |
| Tavris1/ComfyUI-Easy-Install ComfyUI-Easy-Install is a portable one-click installer for ComfyUI that bundles an EZi Desktop application, requiring no manual Python or G… | 85 | 1831 | active |
| TencentARC/BrushNet BrushNet is the official PyTorch implementation of an ECCV 2024 plug-and-play image inpainting model that embeds pixel-level masked image f… | 25 | 1745 | active |
| elixir-nx/bumblebee Bumblebee is an Elixir library providing pre-trained neural network models built on Axon, with integration for downloading models from Hugg… | 89 | 1662 | active |
| Drexubery/ViewCrafter ViewCrafter is a research codebase that uses video diffusion models to synthesize high-fidelity novel views of scenes from a single or spar… | 49 | 1587 | active |
| tin2tin/Pallaidium Pallaidium is a free, open-source generative AI movie studio implemented as a Blender add-on integrated into the Video Sequence Editor (VSE… | 75 | 1520 | active |
| microsoft/Mage Mage is a family of lightweight 4B-parameter multimodal models from Microsoft, including Mage-VL for image and video understanding and Mage… | 57 | 1516 | active |
| NJU-PCALab/STAR STAR is a research implementation of an ICCV 2025 paper performing real-world video super-resolution using spatial-temporal augmentation wi… | 34 | 1495 | active |
| Francis-Rings/StableAnimator StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po… | 41 | 1430 | active |
| wenqsun/DimensionX DimensionX is a research framework that generates photorealistic 3D and 4D scenes from a single image using controllable video diffusion mo… | 43 | 1333 | active |
| shuyu-labs/AntSK AntSK is an enterprise AI knowledge base and agent platform built on .NET 9, Blazor, Semantic Kernel, and Kernel Memory. It supports import… | 53 | 1325 | active |
| W2GenAI-Lab/LucidFlux LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a… | 55 | 1293 | active |
| Tencent-Hunyuan/SRPO SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion… | 54 | 1278 | active |
| rlawjdghek/StableVITON StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent… | 47 | 1261 | stable |
| wladradchenko/wunjo.wladradchenko.ru Wunjo CE is an open-source, locally-run AI media suite for face swap, lip sync, voice cloning, object/text/background removal, restyling, a… | 70 | 1169 | active |
| Woolverine94/biniou biniou is a self-hosted web UI for 30+ generative AI models covering image, video, audio, and text generation, built with Gradio and Huggin… | 63 | 1150 | active |
| PurpleDoubleD/locally-uncensored Locally Uncensored is a free, open-source desktop AI studio (built with Tauri/TypeScript) that bundles uncensored local chat, a coding agen… | 81 | 1146 | active |
| Janspiry/Image-Super-Resolution-via-Iterative-Refinement An unofficial PyTorch implementation of SR3 (Image Super-Resolution via Iterative Refinement), a diffusion-based model for image super-reso… | 32 | 3923 | maintenance |
| OpenMOSS/MOVA MOVA is an open-source foundation model and toolkit for joint video-audio generation, synthesizing synchronized video and audio in a single… | 60 | 1106 | active |
| TencentARC/T2I-Adapter Official implementation of T2I-Adapter, lightweight adapter models that add controllable conditioning (sketch, canny, lineart, depth, pose)… | 31 | 3801 | maintenance |
| JackAILab/ConsistentID ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using … | 52 | 1026 | active |
| MeiGen-AI/PosterCraft PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight… | 48 | 1007 | active |
| tamarott/SinGAN Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat… | 32 | 3344 | maintenance |
| rinongal/textual_inversion Official implementation of the Textual Inversion paper, which learns new word embeddings in a frozen text-to-image (Latent Diffusion) model… | 32 | 3055 | maintenance |
| SCUTlihaoyu/open-chat-video-editor An open-source Python tool that automatically generates short videos from a short text prompt or a web URL, producing narration, background… | 29 | 2813 | maintenance |
| IDEA-Research/DWPose DWPose is the official implementation of 'Effective Whole-body Pose Estimation with Two-stages Distillation' (ICCV 2023), providing whole-b… | 28 | 2807 | maintenance |
| buaacyw/GaussianEditor GaussianEditor is a research application for fast, controllable 3D scene editing built on Gaussian Splatting, released as CVPR 2024 code. I… | 27 | 1455 | maintenance |
| wenhaochai/StableVideo StableVideo is the official ICCV 2023 implementation of text-driven, consistency-aware diffusion-based video editing built on ControlNet an… | 31 | 1439 | maintenance |
| mayuelala/FollowYourPose Official PyTorch implementation of Follow-Your-Pose (AAAI 2024), a pose-guided text-to-video generation model that tunes a text-to-image mo… | 21 | 1358 | maintenance |
| s9roll7/ebsynth_utility An AUTOMATIC1111 WebUI extension that creates stylized or edited videos by combining Stable Diffusion img2img with EbSynth keyframe propaga… | 31 | 1274 | maintenance |
| visual-openllm/visual-openllm An open-source tool that interactively connects different visual models with an LLM, built on ChatGLM, Visual ChatGPT, and Stable Diffusion… | 30 | 1186 | maintenance |
| showlab/Show-1 Show-1 is a research codebase implementing a text-to-video generation model that combines pixel and latent diffusion models, published at I… | 46 | 1148 | maintenance |
| hotshotco/Hotshot-XL Hotshot-XL is an AI text-to-GIF model built to work alongside Stable Diffusion XL, generating 1-second GIFs at 8 FPS. It supports any fine-… | 27 | 1111 | maintenance |
| EdVince/Stable-Diffusion-NCNN A C++ implementation of Stable Diffusion using the NCNN inference framework, supporting both txt2img and img2img. It runs on x86 Windows ex… | 23 | 1067 | maintenance |
| OpenBMB/VisCPM VisCPM is a family of open-source bilingual (Chinese/English) multimodal large models built on the 10B CPM-Bee language model, comprising V… | 29 | 1062 | maintenance |
| pesser/stable-diffusion The development repository for Stable Diffusion and Latent Diffusion Models, containing research code, training scripts, and pretrained mod… | 32 | 1033 | maintenance |
| lllyasviel/LayerDiffuse LayerDiffuse is a research project that generates transparent images and image layers using diffusion models with latent transparency. It p… | 25 | 2221 | experimental |
← prev page 2 / 2