domain: artificial-intelligence
4539 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| EchoMimic EchoMimic is a series of open-source models (V1-V3) from Ant Group for audio-driven human animation, generating lifelike talking-head, port… | 50 | 1038 | active |
| iptag/jimeng-api A self-hosted API service that reverse-engineers Jimeng AI (China) and Dreamina (international) to expose free AI image and video generatio… | 10 | 1038 | active |
| vercel-labs/tersa Tersa is an open-source visual AI playground built as a Next.js web application where users drag, drop, and connect nodes on a canvas to bu… | 63 | 1037 | active |
| Soul-AILab/SoulX-FlashHead SoulX-FlashHead is a 1.3B-parameter framework for high-fidelity, infinite-length, real-time streaming talking-head portrait video generatio… | 53 | 1037 | active |
| webwhiz-ai/webwhiz WebWhiz is an open-source, self-hostable application that trains a ChatGPT-powered chatbot on your website data by crawling your pages and … | 47 | 1037 | active |
| centerforaisafety/HarmBench HarmBench is a standardized, open-source evaluation framework for automated red teaming of large language models, comparing attack methods … | 26 | 1037 | active |
| susiai/susi_chat A chat interface for communicating with a locally hosted LLM (via llama.cpp) through a terminal console, a browser-based console, or a voic… | 60 | 1036 | active |
| ZhangJinHaHaHa/AgentLens AgentLens is a trust-first AI Agent marketplace and workspace where users can discover, compare, and rent task-specific Agents backed by sm… | 58 | 1035 | active |
| inclusionAI/UI-Venus UI-Venus is a family of open-source multimodal GUI agent models (9B/27B) that perform UI element grounding and task navigation from screens… | 63 | 1034 | active |
| villesau/ai-codereviewer A GitHub Action that uses OpenAI's GPT-4 API to automatically review pull requests and post intelligent feedback comments. It retrieves PR … | 21 | 1034 | active |
| mybigday/llama.rn A React Native binding of llama.cpp that enables on-device LLM inference on iOS and Android. It supports GPU/NPU acceleration (Metal, Hexag… | 94 | 1033 | active |
| TTPlanetPig/Comfyui_TTP_Toolset A collection of ComfyUI custom nodes for tiled image processing, including object-aware Smart Tile 2.0 workflows for detail img2img upscali… | 65 | 1033 | active |
| HarborYuan/ovsam Official PyTorch implementation of Open-Vocabulary SAM (ECCV 2024), a model that unifies SAM's interactive segmentation with CLIP's open-vo… | 41 | 1033 | active |
| Agentic Document Extraction (ADE) The official CLI for LandingAI's Agentic Document Extraction (ADE), which parses documents into grounded Markdown and elements and extracts… | 84 | 1032 | active |
| NVIDIA-NeMo/Skills Nemo-Skills is a collection of Python pipelines for improving the skills of large language models, covering synthetic data generation, mode… | 70 | 1032 | active |
| walter-grace/mac-code A free, local coding agent for Apple Silicon Macs that runs large quantized LLMs (e.g., Qwen 35B) via llama.cpp, using techniques like flas… | 48 | 1032 | active |
| bilibili/Index-1.9B Index-1.9B is a family of lightweight 1.9-billion-parameter multilingual language models from Bilibili's Index team, released in base, chat… | 68 | 1031 | active |
| juanjuandog/FinSight-AI FinSight AI is an open-source equity research workspace for A-share companies that turns market data, filings, and financial metrics into s… | 59 | 1031 | active |
| Kotlin/kotlin-agent-skills A collection of AI agent skills for Kotlin projects, packaged as self-contained folders following the Agent Skills standard (agentskills.io… | 59 | 1031 | active |
| andyzoujm/representation-engineering Official library for Representation Engineering (RepE), a top-down approach to AI transparency that monitors and manipulates population-lev… | 18 | 1031 | active |
| Goekdeniz-Guelmez/Local-NotebookLM A local, open-source alternative to Google's NotebookLM that converts PDF documents into audio content like podcasts, summaries, and interv… | 74 | 1030 | active |
| AnimateDiff for Stable Diffusion WebUI A Stable Diffusion WebUI extension that integrates Segment Anything and GroundingDINO to generate segmentation masks from clicks or text pr… | 30 | 3501 | maintenance |
| minimaxir/simpleaichat A Python library providing a minimal, low-complexity interface for building applications on top of chat LLM APIs like ChatGPT and GPT-4. It… | 20 | 3499 | maintenance |
| DLR-RM/3DObjectTracking A collection of C++ implementations of 3D object tracking algorithms from DLR research, including region-based 6DoF trackers (RBGT, SRT3D, … | 49 | 1028 | active |
| agenmod/immortal-skill An open-source 'digital immortality' framework that distills a person's persona from chat logs and documents across 12+ platforms (WeChat, … | 49 | 1028 | active |
| kuprel/min-dalle min(DALL·E) is a fast, minimal PyTorch port of DALL·E Mini/Mega stripped down for text-to-image inference, with only numpy, requests, pillo… | 30 | 3494 | maintenance |
| jianjieyiban/JJYB_AI_VideoAutoCut JJYB_AI 智剪 is a local-first desktop AI video creation workbench that combines material analysis, smart shot segmentation, commentary script… | 70 | 1027 | active |
| Kappaemme-git/codex-first-customer-finder-skill A Codex skill (plugin) that takes a startup URL or product idea and produces an evidence-backed shortlist of potential first customers from… | 57 | 1027 | active |
| 99AI 99AI is a commercially viable, self-hostable AI web platform built with Vue and Node.js that bundles AI chat, image/video/music generation,… | 35 | 1027 | active |
| chrisgoringe/cg-use-everywhere A ComfyUI custom node plugin that provides 'Anything Everywhere' nodes which broadcast data (like MODEL, CLIP, VAE) to all nodes that need … | 69 | 1026 | active |
| mikiarlo3/ai-copywriter A portable Markdown-based agent skill that writes marketing copy (headlines, descriptions, microcopy, subject lines) with a human tone whil… | 55 | 1026 | active |
| JackAILab/ConsistentID ConsistentID is a diffusion-based portrait generation model and toolkit that preserves facial identity from a single reference image using … | 51 | 1026 | active |
| zai-org/GLM-Image GLM-Image is an open-source image generation model combining a 9B autoregressive generator with a 7B diffusion decoder, excelling at text r… | 48 | 1025 | active |
| jucasoliveira/terminalGPT TerminalGPT is a Node.js CLI tool that brings ChatGPT-like LLM chat conversations directly into your terminal. It supports multiple LLM pro… | 36 | 1025 | active |
| MGdaasLab/WHartTest WHartTest is an AI-driven test automation platform built on Django 5.2 + DRF with a Vue frontend, generating and managing executable test c… | 84 | 1024 | active |
| AlexWan/OsEngine OsEngine is an open-source algorithmic trading platform written in C# that bundles a robot-creation layer, historical data downloader (OsDa… | 84 | 1023 | active |
| hrithikkoduri/WebRover WebRover is an autonomous AI web agent that interprets user input, navigates websites via browser automation, and performs tasks or deep re… | 14 | 1023 | active |
| alex-petrenko/sample-factory Sample Factory is a high-throughput Python reinforcement learning library implementing synchronous and asynchronous policy gradient algorit… | 63 | 1022 | active |
| yangxy/PASD PASD (Pixel-Aware Stable Diffusion) is a Python research codebase implementing an ECCV 2024 method for realistic image super-resolution and… | 28 | 1022 | active |
| Tencent/WeSmartFlow WeSmartFlow is an agent-native adaptive learning framework that models the full learning process with ReAct-based tutoring agents, graph me… | 58 | 1021 | active |
| Jingyi-Wu-Richael/rachel-digital-human-production A Codex skill (plugin) that packages a repeatable workflow for producing authorized digital-human talking-head videos using MiniMax voice c… | 54 | 1021 | active |
| Arcade Arcade MCP is an open-source Python framework for building Model Context Protocol (MCP) servers and the tools that run inside them, with a … | 69 | 1020 | active |
| Alpha-VLLM/Lumina-DiMOO Lumina-DiMOO is an open-source omni diffusion large language model that uses fully discrete diffusion to handle multimodal inputs and outpu… | 54 | 1020 | active |
| bxx2004/FUNGA FUNGA (Functional Gene Annotation Exploitation Platform) is an AI-powered web platform for mining gene functions, integrating cross-species… | 38 | 1020 | active |
| Particle Life An interactive desktop application that simulates Particle Life, a simple particle system that exhibits complex life-like emergent behavior… | 62 | 1019 | active |
| google/prompt-to-prompt Google's official implementation of the Prompt-to-Prompt paper, which enables text-driven image editing in Latent Diffusion and Stable Diff… | 10 | 3456 | maintenance |
| Vinyzu/Botright Botright is a Python browser automation framework built on Playwright that provides undetectable, fingerprint-changing stealth browsing. It… | 77 | 1018 | active |
| XGenerationLab/XiYan-SQL XiYan-SQL is a multi-generator ensemble framework for converting natural language questions into SQL queries, achieving SOTA results on ben… | 59 | 1018 | active |
| Towhee Towhee is a Python framework for building ETL pipelines that process unstructured data (images, video, text, audio) into embeddings using s… | 23 | 3453 | maintenance |
| tensorflow/adanet AdaNet is a lightweight TensorFlow-based AutoML framework that automatically learns high-quality neural network architectures and ensembles… | 10 | 3452 | maintenance |
| Sxela/WarpFusion WarpFusion is a Stable Diffusion-based video-to-video style transfer tool distributed as a Jupyter/Colab notebook. It applies AI animation … | 30 | 1016 | active |
| microsoft/aurora Aurora is Microsoft's implementation of a deep learning foundation model for Earth system forecasting, predicting atmospheric variables lik… | 88 | 1015 | active |
| Alpha-VLLM/Lumina-Image-2.0 Lumina-Image 2.0 is an open-source 2.6B-parameter text-to-image generation framework built on a unified Next-DiT architecture with a unifie… | 58 | 1014 | active |
| EvolvingLMMs-Lab/Otter Otter is a multi-modal vision-language model built on OpenFlamingo, instruction-tuned on the MIMIC-IT dataset with image and video understa… | 21 | 3438 | maintenance |
| fallenshock/FlowEdit Official PyTorch implementation of FlowEdit, an ICCV 2025 method for inversion-free, text-based editing of real images using pre-trained fl… | 65 | 1013 | active |
| marcus/nightshift Nightshift is a Go CLI tool that runs overnight, using your leftover Claude or Codex token budget to automatically find issues like dead co… | 70 | 1011 | active |
| RupertAvery/DiffusionToolkit Diffusion Toolkit is a Windows desktop application that indexes and views metadata (prompts, models, settings) embedded in AI-generated ima… | 64 | 1010 | active |
| thuml/Large-Time-Series-Model Official code, datasets, and checkpoints for Timer and Sundial, generative pre-trained Transformer foundation models for general time serie… | 58 | 1010 | active |
| HumeAI/tada TADA is an open-source speech-language model from Hume AI that generates expressive, high-fidelity speech via text-acoustic dual alignment,… | 51 | 1010 | active |
| 274056675/springboot-openai-chatgpt A full-stack AI chatbot application built on Spring Boot/Spring Cloud that integrates GPT-3.5, GPT-4, Baidu ERNIE Bot, Stable Diffusion, an… | 36 | 1010 | active |
| sail-sg/EditAnything Edit Anything is a Python application for text-guided image editing and generation, combining Segment Anything, ControlNet, BLIP2, and Stab… | 33 | 3422 | maintenance |
| octimot/StoryToolkitAI StoryToolkitAI is a desktop film editing tool that transcribes, indexes, and semantically searches video footage locally, using speech reco… | 65 | 1009 | active |
| MeiGen-AI/PosterCraft PosterCraft is a unified framework for generating high-quality aesthetic posters, published as an ICLR 2026 paper. It provides model weight… | 48 | 1009 | active |
| 199-biotechnologies/claude-deep-research-skill A Claude Code skill (plugin) that turns the agent into an enterprise-grade deep research engine with an 8-phase pipeline covering scoping, … | 51 | 1008 | active |
| ZTE-AICloud/Co-Sight Co-Sight is an open-source Python framework from ZTE for building Manus-like autonomous AI agent systems that generate high-quality researc… | 44 | 1008 | active |
| Kosinkadink/ComfyUI-Advanced-ControlNet A set of ComfyUI custom nodes providing advanced ControlNet scheduling, weighting, and masking for Stable Diffusion workflows. It supports … | 71 | 1007 | active |
| MATLAB Agentic Toolkit Ecosystem A MATLAB toolkit that connects AI coding agents to MATLAB by installing the MATLAB MCP Server and providing curated skills with MATLAB work… | 81 | 1006 | active |
| voquill/voquill Voquill is an open-source, cross-platform AI voice dictation app that lets users dictate into any desktop application, with AI-powered tran… | 76 | 1006 | active |
| gausian-AI/Gausian_native_editor Gausian is a native desktop video editor built in Rust with GPU-accelerated preview (WGPU), timeline editing, and hardware decoding via Vid… | 52 | 1006 | active |
| alumnium-hq/alumnium Alumnium is an AI-native library and MCP server for end-to-end testing that layers natural-language actions, checks, and data extraction on… | 90 | 1005 | active |
| huawei-noah/noah-research A collection of research code subprojects released by Huawei Noah's Ark Lab, each in its own directory. It is not an official Huawei produc… | 76 | 1005 | active |
| vercel-labs/knowledge-agent-template An open-source TypeScript (Nuxt/Vue) template for building file-system and knowledge-base based AI agents that search sources with grep, fi… | 60 | 1005 | active |
| FastEmbed A Rust library for generating text and image vector embeddings and reranking documents locally using ONNX inference via ort and Hugging Fac… | 94 | 1004 | active |
| shaoshengsong/DeepSORT A real-time multi-object tracking application in C++17 using Qt 6, OpenCV, and ONNX Runtime, running YOLO detection with four pluggable tra… | 89 | 1004 | active |
| ZJUI-AI4H/Hulu-Med Hulu-Med is a family of open-source transparent generalist medical vision-language models ranging from 4B to 235B parameters, covering text… | 61 | 1004 | active |
| InsiderX-Pro/video-translator An OpenClaw agent skill (written in Shell/Python) that translates and dubs videos by submitting jobs to a remote video-translation service … | 54 | 1004 | active |
| MichaelYuhe/ai-group-tabs A Chrome extension that uses AI (OpenAI API) to automatically organize and group browser tabs into categories. Users can customize categori… | 10 | 1004 | active |
| Overcooked AI Overcooked-AI is a benchmark environment for fully cooperative human-AI task performance, based on the video game Overcooked, where agents … | 28 | 1003 | active |
| QwenAudio/Fun-Audio-Chat Fun-Audio-Chat is a Large Audio Language Model (8B) for natural, low-latency voice interactions, using dual-resolution speech representatio… | 46 | 1002 | active |
| WangRongsheng/CareGPT CareGPT is a medical large language model project that aggregates dozens of publicly available medical fine-tuning datasets and open medica… | 10 | 1002 | active |
| jonexaiorg/jonex Jonex is an end-to-end, self-hosted AI knowledge platform combining a multimodal parsing engine (documents, video, images) with an ontology… | 58 | 1001 | active |
| siyuanchen0214/Scam-AI-Multi-modal-Evaluation-System A Python-based multi-modal AI system for detecting fraudulent content across text, image, audio, and video, with provenance tracing and cro… | 38 | 1001 | active |
| argilla-io/distilabel Distilabel is a Python framework for building scalable pipelines that generate synthetic data and AI feedback, based on verified research p… | 74 | 3385 | maintenance |
| talkie-lm/talkie talkie is a Python library and CLI for running inference with the talkie 13B language model family, including a 13B model trained on pre-19… | 51 | 1000 | active |
| disler/always-on-ai-assistant A Python reference pattern for an always-on voice AI assistant that listens via RealtimeSTT, thinks with Deepseek-V3 or local Ollama models… | 21 | 998 | active |
| davidrmiller/biosim4 A C++ command-line program that simulates biological creatures evolving through natural selection in a 2D arena, as featured in the YouTube… | 71 | 3364 | maintenance |
| LLM-As-Chatbot A Python application that serves open-source instruction-tuned LLMs as chatbot services via a Gradio web UI or a Discord bot. It integrates… | 31 | 3319 | maintenance |
| X-D-Lab/LangChain-ChatGLM-Webui A Gradio-based web UI that combines LangChain with ChatGLM-6B and other open-source LLMs to answer questions over a local knowledge base. U… | 30 | 3312 | maintenance |
| PixArt-alpha/PixArt-alpha PixArt-α is a Transformer-based text-to-image diffusion model with PyTorch model definitions, pre-trained weights, and inference/training c… | 27 | 3304 | maintenance |
| resemble-ai/Resemblyzer Resemblyzer is a Python package that uses a deep learning voice encoder to convert speech audio into 256-dimensional voice embeddings. Thes… | 23 | 3302 | maintenance |
| martinarjovsky/WassersteinGAN Reference PyTorch implementation of the Wasserstein GAN paper, providing training scripts for DCGAN and MLP architectures on datasets like … | 32 | 3244 | maintenance |
| google/model_search Model Search is a Google AutoML framework that implements neural architecture search algorithms at scale to find optimal DNN architectures … | 10 | 3238 | maintenance |
| FreeGPT35 A self-hosted service that exposes the login-free ChatGPT Web's GPT-3.5-Turbo as an OpenAI-compatible /v1/chat/completions API. It can be d… | 10 | 3228 | maintenance |
| JoePenna/Dreambooth-Stable-Diffusion A Jupyter Notebook-based implementation of Dreambooth fine-tuning for Stable Diffusion, adapted from XavierXiao's repo with tweaks for trai… | 32 | 3212 | maintenance |
| VinsonLaro/stable-diffusion-webui-chinese A Simplified Chinese localization extension for the AUTOMATIC1111 Stable Diffusion WebUI, provided as JSON translation templates. It also i… | 32 | 3205 | maintenance |
| pythongosssss/ComfyUI-Custom-Scripts A ComfyUI custom node extension providing UI enhancements and quality-of-life features such as prompt autocomplete, graph auto-arrangement,… | 60 | 3177 | maintenance |
| project-baize/baize-chatbot Baize is an open-source chat model built on LLaMA using LoRA parameter-efficient tuning, trained on 100k self-chat dialogs generated by Cha… | 21 | 3148 | maintenance |
| DAMO-NLP-SG/Video-LLaMA Video-LLaMA is an instruction-tuned audio-visual language model that extends LLaMA with video and audio understanding via cross-modal pretr… | 30 | 3141 | maintenance |
| Habitat AI Habitat is a high-performance, physics-enabled 3D simulation platform for Embodied AI research, consisting of Habitat-Sim (a fast 3D sim… | 74 | 3118 | maintenance |
| salesforce/CodeT5 Official research release of CodeT5 and CodeT5+ open code large language models from Salesforce Research for code understanding and generat… | 10 | 3093 | maintenance |