Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: stable-diffusion

161 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
ControlNet
ControlNet is a neural network architecture that adds conditional control (edges, poses, depth, etc.) to pretrained text-to-image diffusion…
3134091maintenance
ddean2009/MoneyPrinterPlus
A Python desktop application that uses AI LLMs to batch-generate short videos with one click, including automatic video mashup/remixing and…
307008active
leejet/stable-diffusion.cpp
A pure C/C++ inference engine for diffusion models (Stable Diffusion, FLUX, Wan, Qwen Image, Z-Image, and more) built on ggml in the style …
916846active
Doubiiu/ToonCrafter
ToonCrafter is a generative model that interpolates two cartoon images into a short animation by leveraging pre-trained image-to-video diff…
296003stable
aidlearning/AidLearning-FrameWork
AidLux (originally AidLearning) is an AIoT development platform that runs a native Ubuntu Linux environment with GUI, deep learning tooling…
705797active
xxlong0/Wonder3D
Wonder3D is a cross-domain diffusion model that reconstructs high-fidelity textured 3D meshes from a single image in 2-3 minutes. It genera…
325425active
yisol/IDM-VTON
Official implementation of IDM-VTON, an ECCV 2024 paper that improves diffusion models for high-fidelity virtual try-on, swapping garments …
305156active
tyxsspa/AnyText
AnyText is the official implementation of a diffusion-based model for multilingual visual text generation and editing in images, accepted a…
324874active
xlite-dev/lite.ai.toolkit
A lightweight C++ toolkit providing unified APIs for 100+ pre-trained AI models across inference backends like ONNX Runtime, MNN, TensorRT,…
744427active
dreamgaussian/dreamgaussian
DreamGaussian is the official PyTorch implementation of an ICLR 2024 Oral paper for efficient 3D content creation using generative Gaussian…
184352active
IDEA-CCNL/Fengshenbang-LM
Fengshenbang-LM is an open-source suite of Chinese large language models and a PyTorch training framework from IDEA Research's CCNL lab, ai…
714123active
Kosinkadink/ComfyUI-AnimateDiff-Evolved
A ComfyUI custom node pack providing an improved AnimateDiff integration plus advanced 'Evolved Sampling' options for animated video genera…
703531active
MooreThreads/Moore-AnimateAnyone
An open-source reproduction of AnimateAnyone that animates a character from a single reference image using pose sequences from a driving vi…
263514active
CompVis/latent-diffusion
The official research code and pretrained model zoo for Latent Diffusion Models (LDM), the paper behind Stable Diffusion, enabling high-res…
3214133maintenance
dirk1983/deepseek
A lightweight PHP demo application that calls DeepSeek/OpenAI-compatible chat APIs with streaming (EventSource) output, supporting Markdown…
363103active
sonos/tract
Tract is Sonos' tiny, self-contained neural-network inference engine written in Rust. It loads ONNX, TensorFlow/TFLite, and NNEF models, op…
993045active
off-grid-ai/OGAM
Off Grid AI (OGAM) is a cross-platform mobile and desktop application that runs AI entirely on-device: GGUF LLM chat with vision, Whisper s…
783002active
TMElyralab/MuseV
MuseV is a diffusion-based framework for generating high-fidelity virtual human videos of infinite length using a Visual Conditioned Parall…
252846active
bytedance/InfiniteYou
InfiniteYou (InfU) is a research framework from ByteDance for identity-preserved text-to-image generation built on Diffusion Transformers l…
372685active
IceClear/StableSR
StableSR is a Python research library that leverages pre-trained Stable Diffusion priors for real-world blind image super-resolution. It pr…
212668stable
koishijs/novelai-bot
A Koishi chatbot plugin that generates images via NovelAI, with support for SD-WebUI and Stable Horde backends. It offers model/sampler/siz…
362550active
yifan123/flow_grpo
Flow-GRPO is the official PyTorch implementation of a NeurIPS 2025 paper that trains flow matching models (e.g., SD3.5, FLUX.1, Qwen-Image,…
562498active
learnhouse/learnhouse
LearnHouse is a next-generation open-source learning management system (LMS) for creating, sharing, and selling educational content. It com…
992203active
SUDO-AI-3D/zero123plus
Zero123++ is a diffusion base model that generates consistent multi-view images from a single input image, intended as a stepping stone for…
272095active
storytold/artcraft
ArtCraft is an open-source desktop application for interactive AI image and video creation, described as 'the IDE for artists'. It provides…
942044active
Tavris1/ComfyUI-Easy-Install
ComfyUI-Easy-Install is a portable one-click installer for ComfyUI that bundles an EZi Desktop application, requiring no manual Python or G…
851831active
TencentARC/BrushNet
BrushNet is the official PyTorch implementation of an ECCV 2024 plug-and-play image inpainting model that embeds pixel-level masked image f…
251745active
elixir-nx/bumblebee
Bumblebee is an Elixir library providing pre-trained neural network models built on Axon, with integration for downloading models from Hugg…
891662active
Drexubery/ViewCrafter
ViewCrafter is a research codebase that uses video diffusion models to synthesize high-fidelity novel views of scenes from a single or spar…
491587active
tin2tin/Pallaidium
Pallaidium is a free, open-source generative AI movie studio implemented as a Blender add-on integrated into the Video Sequence Editor (VSE…
751520active
microsoft/Mage
Mage is a family of lightweight 4B-parameter multimodal models from Microsoft, including Mage-VL for image and video understanding and Mage…
571516active
NJU-PCALab/STAR
STAR is a research implementation of an ICCV 2025 paper performing real-world video super-resolution using spatial-temporal augmentation wi…
341495active
Francis-Rings/StableAnimator
StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po…
411430active
wenqsun/DimensionX
DimensionX is a research framework that generates photorealistic 3D and 4D scenes from a single image using controllable video diffusion mo…
431333active
shuyu-labs/AntSK
AntSK is an enterprise AI knowledge base and agent platform built on .NET 9, Blazor, Semantic Kernel, and Kernel Memory. It supports import…
531325active
W2GenAI-Lab/LucidFlux
LucidFlux is a caption-free photo-realistic image restoration model built on a large-scale diffusion transformer, released with inference a…
551293active
Tencent-Hunyuan/SRPO
SRPO is Tencent Hunyuan's research code for fine-tuning diffusion image generation models (e.g., FLUX.1.dev) by aligning the full diffusion…
541278active
rlawjdghek/StableVITON
StableVITON is the official PyTorch implementation of a CVPR 2024 paper that performs image-based virtual try-on using a pre-trained latent…
471261stable
wladradchenko/wunjo.wladradchenko.ru
Wunjo CE is an open-source, locally-run AI media suite for face swap, lip sync, voice cloning, object/text/background removal, restyling, a…
701169active
Woolverine94/biniou
biniou is a self-hosted web UI for 30+ generative AI models covering image, video, audio, and text generation, built with Gradio and Huggin…
631150active
PurpleDoubleD/locally-uncensored
Locally Uncensored is a free, open-source desktop AI studio (built with Tauri/TypeScript) that bundles uncensored local chat, a coding agen…
811146active
Janspiry/Image-Super-Resolution-via-Iterative-Refinement
An unofficial PyTorch implementation of SR3 (Image Super-Resolution via Iterative Refinement), a diffusion-based model for image super-reso…
323923maintenance
OpenMOSS/MOVA
MOVA is an open-source foundation model and toolkit for joint video-audio generation, synthesizing synchronized video and audio in a single…
601106active
TencentARC/T2I-Adapter
Official implementation of T2I-Adapter, lightweight adapter models that add controllable conditioning (sketch, canny, lineart, depth, pose)…
313801maintenance
JackAILab/ConsistentID
ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using …
521026active
MeiGen-AI/PosterCraft
PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight…
481007active
tamarott/SinGAN
Official PyTorch implementation of SinGAN, an ICCV 2019 best-paper generative model trained on a single natural image. It learns patch stat…
323344maintenance
rinongal/textual_inversion
Official implementation of the Textual Inversion paper, which learns new word embeddings in a frozen text-to-image (Latent Diffusion) model…
323055maintenance
SCUTlihaoyu/open-chat-video-editor
An open-source Python tool that automatically generates short videos from a short text prompt or a web URL, producing narration, background…
292813maintenance
IDEA-Research/DWPose
DWPose is the official implementation of 'Effective Whole-body Pose Estimation with Two-stages Distillation' (ICCV 2023), providing whole-b…
282807maintenance
buaacyw/GaussianEditor
GaussianEditor is a research application for fast, controllable 3D scene editing built on Gaussian Splatting, released as CVPR 2024 code. I…
271455maintenance
wenhaochai/StableVideo
StableVideo is the official ICCV 2023 implementation of text-driven, consistency-aware diffusion-based video editing built on ControlNet an…
311439maintenance
mayuelala/FollowYourPose
Official PyTorch implementation of Follow-Your-Pose (AAAI 2024), a pose-guided text-to-video generation model that tunes a text-to-image mo…
211358maintenance
s9roll7/ebsynth_utility
An AUTOMATIC1111 WebUI extension that creates stylized or edited videos by combining Stable Diffusion img2img with EbSynth keyframe propaga…
311274maintenance
visual-openllm/visual-openllm
An open-source tool that interactively connects different visual models with an LLM, built on ChatGLM, Visual ChatGPT, and Stable Diffusion…
301186maintenance
showlab/Show-1
Show-1 is a research codebase implementing a text-to-video generation model that combines pixel and latent diffusion models, published at I…
461148maintenance
hotshotco/Hotshot-XL
Hotshot-XL is an AI text-to-GIF model built to work alongside Stable Diffusion XL, generating 1-second GIFs at 8 FPS. It supports any fine-…
271111maintenance
EdVince/Stable-Diffusion-NCNN
A C++ implementation of Stable Diffusion using the NCNN inference framework, supporting both txt2img and img2img. It runs on x86 Windows ex…
231067maintenance
OpenBMB/VisCPM
VisCPM is a family of open-source bilingual (Chinese/English) multimodal large models built on the 10B CPM-Bee language model, comprising V…
291062maintenance
pesser/stable-diffusion
The development repository for Stable Diffusion and Latent Diffusion Models, containing research code, training scripts, and pretrained mod…
321033maintenance
lllyasviel/LayerDiffuse
LayerDiffuse is a research project that generates transparent images and image layers using diffusion models with latent transparency. It p…
252221experimental

← prev page 2 / 2