domain: artificial-intelligence
4539 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| aiming-lab/Agent0 Agent0 Series is a research framework for training self-evolving LLM agents from zero external data via tool-integrated reasoning and co-ev… | 58 | 1258 | active |
| bug0inc/passmark Passmark is an open-source TypeScript library that extends Playwright to run end-to-end browser regression tests written in plain English. … | 57 | 1258 | active |
| Francis-Rings/StableAvatar StableAvatar is an end-to-end video diffusion transformer that generates infinite-length, high-quality talking avatar videos from a referen… | 46 | 1258 | active |
| najmuzzaman-mohammad/gawkbot WUPHF (gawkbot) is a locally-run application that turns described manual workflows into AI agent 'microapps' with their own UI, schedules, … | 78 | 1257 | active |
| FaceAISDK An Android SDK for fully on-device, offline face detection, recognition, liveness detection (anti-spoofing), and 1:1, 1:N, and M:N face sea… | 98 | 1256 | active |
| MiniMax-AI/OpenRoom VibeApps (OpenRoom) is a browser-based desktop environment with draggable windows and a suite of built-in apps, all controlled by an AI Age… | 53 | 1256 | active |
| FoloUp/FoloUp FoloUp is an open-source, self-hostable web application that conducts AI-powered voice interviews with job candidates. It generates tailore… | 58 | 1255 | active |
| myreader-io/myGPTReader myGPTReader is a Slack bot powered by ChatGPT that reads and summarizes webpages, documents (eBooks, PDF, DOCX), and YouTube videos, and su… | 60 | 4417 | maintenance |
| minsight-ai-info/AI-Search-Hub AI Search Hub is an open-source Skill that aggregates native AI search capabilities from platforms like Gemini, Grok, Doubao, and Yuanbao i… | 50 | 1254 | active |
| tigicion/dao-code Dao Code is an open-source terminal-native AI coding agent written in TypeScript, targeting DeepSeek V4 with 1M context and Claude Code-com… | 76 | 1253 | active |
| zjunlp/SkillNet SkillNet is open infrastructure for discovering, evaluating, composing, and orchestrating reusable AI agent skills, treating skills as sear… | 59 | 1253 | active |
| BehiSecc/VibeSec-Skill VibeSec-Skill is an AI skill (prompt/instruction pack) that teaches LLM coding assistants like Claude Code, Cursor, Codex, Copilot, and Ant… | 45 | 1253 | active |
| Anthony's QR Code Toolkit A web-based toolkit for generating base QR codes and refining AI-generated QR codes by comparing outputs to find misaligned pixels. It also… | 29 | 1253 | active |
| tsinghua-fib-lab/AgentSociety AgentSociety is an LLM-native agent simulation framework for building large-scale multi-agent simulations of social and urban environments.… | 86 | 1252 | active |
| X-Square-Robot/wall-x Wall-X is the open-source training and inference stack for X Square Robot's WALL series of embodied foundation models (VLAs) for general-pu… | 62 | 1252 | active |
| Linketic/CityGaussian Official implementation of the CityGaussian series (ECCV 2024, ICLR 2025) for high-quality large-scale 3D scene reconstruction with Gaussia… | 66 | 1251 | active |
| ksimback/hermes-ecosystem Hermes Atlas is a community-curated web directory mapping the ecosystem of tools, skills, plugins, and integrations built around Hermes Age… | 59 | 1250 | active |
| JoyCaption JoyCaption is an open, free, and uncensored image captioning Visual Language Model (VLM) with released weights and training scripts. It gen… | 53 | 1250 | active |
| lpiccinelli-eth/UniDepth UniDepth is a Python library and research codebase for universal monocular metric depth estimation from single images, based on CVPR 2024 a… | 34 | 1250 | active |
| m5stack/StackChan StackChan is an open-source kawaii AI desktop robot co-created by M5Stack and its community, built around the CoreS3 (ESP32-S3) IoT develop… | 59 | 1249 | active |
| Roblox/cube Cube is Roblox's open-source family of foundation models for 3D intelligence, including text-to-3D shape generation and part-controllable m… | 57 | 1248 | active |
| unum-cloud/UForm UForm is a compact multimodal AI library providing tiny image-text embedding models (64-768 dimensions, Matryoshka-style) and small generat… | 54 | 1248 | active |
| XPixelGroup/HYPIR Official PyTorch implementation of HYPIR, a SIGGRAPH 2025 method that harnesses diffusion-yielded score priors for image restoration. It pr… | 39 | 1248 | active |
| digitalinnovationone/dio-agent DIO Agent is an AI study mentor created by Digital Innovation One (DIO) that guides learners through bootcamps, courses, and coding challen… | 55 | 1247 | active |
| Heavrnl/TelegramForwarder A self-hosted Telegram message forwarder built on Telethon that copies messages from multiple source chats to target chats with keyword/reg… | 52 | 1247 | active |
| Audio Webui An all-in-one web UI for audio-related neural networks, bundling text-to-speech (Bark), voice conversion/cloning (RVC), and text-to-audio/m… | 30 | 1246 | active |
| 3DTopia/OpenLRM OpenLRM is an open-source PyTorch implementation of Large Reconstruction Models (LRM) that reconstruct 3D objects (meshes and rendered vide… | 17 | 1246 | active |
| wfjsw/danbooru-diffusion-prompt-builder A web application ('Danbooru Tag Supermarket') for browsing, searching, and composing Danbooru/NovelAI tag prompts for Stable Diffusion ima… | 32 | 1244 | active |
| automl/SMAC3 SMAC3 is a Python library for Bayesian Optimization used to tune hyperparameters of machine learning algorithms and configure arbitrary alg… | 81 | 1243 | active |
| thetahealth/mirobody Mirobody is an open-source, AI-native health data engine that collects readings from lab reports, wearables, and genomics, standardizes the… | 84 | 1242 | active |
| pyang5166/gbro-collage-broll An agent skill that turns short voiceover lines into editorial halftone paper-collage B-roll videos using Gemini Omni Flash first/last-fram… | 54 | 1242 | active |
| showlab/Tune-A-Video Tune-A-Video is the official PyTorch implementation of an ICCV 2023 paper that fine-tunes pre-trained text-to-image diffusion models (like … | 31 | 4363 | maintenance |
| alibaba/Tora Tora is Alibaba's official implementation of a trajectory-oriented Diffusion Transformer (DiT) for controllable video generation, integrati… | 64 | 1241 | active |
| JuneYaooo/gpt-image2-ppt-skills A Claude Code / OpenClaw skill that generates polished 16:9 presentations using OpenAI's gpt-image-2, rendering each slide as a complete vi… | 58 | 1241 | active |
| 0xacx/chatGPT-shell-cli A lightweight shell script that lets you chat with OpenAI's ChatGPT models and generate DALL-E images directly from the terminal, requiring… | 32 | 1241 | active |
| google-deepmind/android_env AndroidEnv is a Python library from DeepMind that exposes an Android device (real or emulated) as a Reinforcement Learning environment. Age… | 84 | 1240 | active |
| mrwadams/attackgen AttackGen is a Streamlit-based cybersecurity tool that uses large language models to generate tailored incident response testing scenarios … | 95 | 1239 | active |
| xllm-go/bypass A Go server that reverse-engineers the chat interfaces of multiple AI providers (Coze, DeepSeek, Cursor, Windsurf, Grok, Bing Copilot, You,… | 62 | 1238 | active |
| declare-lab/tango Tango is a family of latent diffusion models for text-to-audio generation, with Tango 2 improving prompt alignment via DPO-based fine-tunin… | 45 | 1238 | active |
| sh-lee-prml/HierSpeechpp Official PyTorch implementation of HierSpeech++, a fast zero-shot speech synthesizer for text-to-speech and voice conversion based on hiera… | 28 | 1238 | active |
| Ksuriuri/index-tts-vllm A reimplementation of IndexTTS's GPT model inference using vLLM, providing significantly faster text-to-speech generation with a web UI and… | 57 | 1237 | active |
| supavec/supavec Supavec is an open-source RAG-as-a-Service platform (an alternative to Carbon.ai) that lets developers upload documents, generate embedding… | 47 | 1236 | active |
| bytedance/USO USO is ByteDance's open-source unified style- and subject-driven image generation model based on diffusion (FLUX), combining any subject wi… | 36 | 1236 | active |
| GreatScottyMac/RooFlow RooFlow is an experimental set of YAML-based system prompts and modes for the Roo Code VS Code extension, providing persistent project cont… | 29 | 1235 | active |
| chengzeyi/Comfy-WaveSpeed A ComfyUI custom node plugin that acts as an all-in-one inference optimization solution for diffusion models, built around First Block Cach… | 66 | 1234 | stable |
| florestefano1975/comfyui-portrait-master A ComfyUI custom node suite that helps AI image creators generate detailed, professional prompts for human portraits. It provides modular n… | 56 | 1233 | active |
| kellyvv/PhoneClaw PhoneClaw is a mobile-native local AI agent framework that turns phones into on-device agent runtimes, running Gemma models via LiteRT and … | 76 | 1232 | active |
| Maia Chess Maia is a collection of human-like neural network chess engines trained on millions of human games, targeting skill levels from ELO 1100 to… | 60 | 1232 | active |
| TheSmallHanCat/sora2api A self-hosted OpenAI-compatible API gateway that wraps Sora's text-to-video and image generation capabilities behind standard /v1/chat/comp… | 10 | 1232 | active |
| ipa-lab/hackingBuddyGPT HackingBuddyGPT is a Python framework that helps ethical hackers and security researchers use LLMs and LLM-based autonomous agents for pene… | 68 | 1230 | active |
| pythongosssss/ComfyUI-WD14-Tagger A ComfyUI custom node extension that interrogates images to extract booru-style tags using WD 1.4 tagger models (ONNX-based). It integrates… | 43 | 1230 | active |
| lucidrains/deep-daze Deep Daze is a simple command line tool for text-to-image generation that combines OpenAI's CLIP with a Siren implicit neural representatio… | 23 | 4313 | maintenance |
| Scale3-Labs/langtrace Langtrace is an open-source, OpenTelemetry-based observability platform for LLM applications, providing real-time tracing, metrics, and eva… | 50 | 1229 | active |
| AWorld AWorld is an open-source agent framework and runtime that orchestrates tools, memory, context, and execution for building autonomous AI age… | 78 | 1228 | active |
| OStudi/short-video-generator-AI An open-source Python tool that turns YouTube videos into ready-to-post vertical short videos by automatically detecting highlights, adding… | 57 | 1228 | active |
| rohunvora/x-research-skill A TypeScript CLI and agent skill that wraps the X/Twitter API for searching tweets, pulling threads, monitoring accounts, and generating so… | 45 | 1228 | active |
| Tencent-Hunyuan/HunyuanCustom HunyuanCustom is a multimodal-driven customized video generation framework built on HunyuanVideo, supporting image, text, audio, and video … | 40 | 1228 | active |
| bricks-cloud/BricksLLM BricksLLM is a cloud-native AI gateway written in Go that sits between applications and LLM providers like OpenAI, Anthropic, Azure OpenAI,… | 30 | 1228 | active |
| apple/python-apple-fm-sdk Python bindings for Apple's Foundation Models framework, giving access to the on-device foundation model behind Apple Intelligence on macOS… | 75 | 1226 | active |
| iDC-NEU/YiGraph YiGraph is an LLM-driven agent system for autonomous graph data analytics built on the Analytics-Augmented Generation (AAG) framework. It e… | 60 | 1226 | active |
| MoonshotAI/Kimi-VL Kimi-VL is an open-source Mixture-of-Experts vision-language model (VLM) with a 2.8B activated parameter language decoder, offering multimo… | 33 | 1226 | active |
| benrugg/AI-Render A Blender add-on that renders AI-generated images with Stable Diffusion based on a text prompt and the user's 3D scene. It supports cloud r… | 57 | 1225 | active |
| Bolin97/GongBU GongBU is a self-hosted, no-code web platform for fine-tuning, evaluating, and deploying large language models, built on Transformers and P… | 52 | 1225 | active |
| mcmonkeyprojects/sd-dynamic-thresholding A Stable Diffusion extension that enables using higher CFG scale values without color artifacts by clamping latents between sampling steps.… | 36 | 1225 | active |
| ElectricAlexis/NotaGen NotaGen is a symbolic music generation model that produces high-quality classical sheet music using LLM-style training paradigms: pre-train… | 31 | 1225 | active |
| run-house/kubetorch Kubetorch is a Python library that lets you distribute and run ML workloads (training, inference, data processing) on Kubernetes directly f… | 82 | 1224 | active |
| sums001/Windows-Copilot-API A Python library and local server that reverse engineers Microsoft Copilot's free web chat into an OpenAI-compatible REST API, giving acces… | 53 | 1224 | active |
| MaoXiaoYuZ/Long-Novel-GPT Long-Novel-GPT is a self-hosted web application that uses LLMs and RAG to generate and revise long-form novels through a top-down outline-c… | 48 | 1224 | active |
| bowang-lab/MedRAX MedRAX is a medical reasoning agent framework that integrates chest X-ray analysis tools (segmentation, grounding, report generation, disea… | 42 | 1223 | active |
| 666ghj/DeepSearchAgent-Demo A framework-free Python implementation of a deep search AI agent that generates high-quality research reports through multi-round web searc… | 34 | 1223 | active |
| allwefantasy/auto-coder Auto-Coder is an open-source AI coding agent and CLI tool (powered by Byzer-LLM) that provides chat, one-shot command, server, and RAG mode… | 69 | 1222 | active |
| GML-MMGroup/GMTalker GMTalker is an interactive 3D digital human system rendered with Unreal Engine, integrating speech recognition, speech synthesis, natural l… | 45 | 1222 | active |
| AllAboutAI-YT/easy-local-rag A Python application providing 100% local retrieval-augmented generation (RAG) using Ollama for LLM inference and embeddings. It supports i… | 15 | 1222 | active |
| jieyefriic/rp-engine A YAML-native agent workflow execution engine written in Rust that parses declarative workflow files describing nodes, edges, prompts, data… | 51 | 1221 | active |
| zkonduit/ezkl EZKL is a Rust-based library and command-line tool that converts deep learning models and arbitrary computational graphs (exported as ONNX)… | 75 | 1220 | active |
| SimonSchubert/Kai Kai 9000 is an open-source, cross-platform AI assistant built with Kotlin Multiplatform that runs on Android, iOS, Windows, Mac, Linux, and… | 84 | 1219 | active |
| microsoft/malmo Project Malmo is a platform for artificial intelligence experimentation and research built on top of Minecraft, providing a gym-like enviro… | 10 | 4270 | maintenance |
| uezo/ChatdollKit ChatdollKit is a Unity SDK that turns 3D character models into voice-enabled chatbots and virtual assistants. It integrates LLMs (ChatGPT, … | 76 | 1218 | active |
| lucidrains/perceiver-pytorch A PyTorch implementation of the Perceiver architecture (General Perception with Iterative Attention) and its follow-up Perceiver IO. It pro… | 61 | 1217 | active |
| fudan-generative-vision/champ Champ is a research framework for controllable and consistent human image animation using 3D parametric guidance (SMPL-based depth, normal,… | 25 | 4263 | maintenance |
| nneonneo/2048-ai An AI solver for the 2048 game using expectimax optimization with an efficient bitboard representation, searching over 10 million moves per… | 65 | 1216 | stable |
| wanxingai/LightAgent LightAgent is an ultra-lightweight open-source Python framework for building LLM-powered agents with tools, persistent memory, guardrails, … | 86 | 1215 | active |
| TIGER-AI-Lab/OpenResearcher OpenResearcher is a fully open-source pipeline for synthesizing long-horizon deep research trajectories using LLM agents with retrieval and… | 54 | 1215 | active |
| tonyqinatcmu/SlideBot-AI SlideBot AI is an AI-powered presentation generator that turns a topic, outline, or uploaded materials (documents, spreadsheets, meeting re… | 44 | 1215 | active |
| Shopify/roast Roast is a Ruby-based domain-specific language for building structured AI workflows from composable building blocks called cogs. It orchest… | 82 | 1213 | active |
| mbrg/power-pwn Power Pwn is an offensive and defensive security toolset for Microsoft 365 Power Platform and AI services, including Copilot Studio, custom… | 62 | 1213 | active |
| ifzhang/FairMOT FairMOT is a research implementation of a one-shot multi-object tracking model that jointly performs object detection and re-identification… | 32 | 4245 | maintenance |
| Picsart-AI-Research/Text2Video-Zero Official implementation of Text2Video-Zero, a zero-shot text-to-video generation method that adapts text-to-image diffusion models like Sta… | 30 | 4243 | maintenance |
| modal-labs/quillman QuiLLMan is a voice chat application built on Kyutai's Moshi speech-to-speech language model, deployed serverlessly on Modal with a FastAPI… | 67 | 1212 | active |
| alibaba/lumenx LumenX is an AI-native platform for turning novel text into publishable motion comic and short drama videos. It provides a full pipeline fr… | 58 | 1209 | active |
| taranis-ai/taranis-ai Taranis AI is a self-hosted open-source OSINT platform that collects news articles from web sources and uses NLP/AI to enrich, cluster, and… | 94 | 1208 | active |
| ardha27/AI-Song-Cover-RVC A collection of Google Colab and Kaggle notebooks that form an all-in-one toolkit for creating AI song covers with RVC (Retrieval-based Voi… | 67 | 1208 | active |
| sashiko-dev/sashiko Sashiko is an agentic code review system for the Linux kernel that ingests patches from mailing lists, GitHub PRs, GitLab MRs, or local git… | 61 | 1208 | active |
| SalesforceAIResearch/enterprise-deep-research Enterprise Deep Research (EDR) is a multi-agent deep research system from Salesforce AI Research that combines a master planning agent, spe… | 55 | 1205 | active |
| hkjarral/AVA-AI-Voice-Agent-for-Asterisk An open-source AI voice agent that integrates with Asterisk/FreePBX phone systems via Audiosocket/RTP, built in Python with a modular pipel… | 84 | 1204 | active |
| modelscope/sirchmunk Sirchmunk is an agentic, embedding-free search engine that turns raw files into a self-evolving knowledge base in real time, without vector… | 82 | 1204 | active |
| NyxTides/ppt-image-first A conversation-first, image-first PPT workflow skill for AI coding CLIs (Codex, Claude Code, Opencode) that turns vague presentation reques… | 50 | 1203 | active |
| metavoiceio/metavoice-src MetaVoice-1B is a 1.2B parameter foundational text-to-speech model trained on 100K hours of speech, focused on emotional rhythm and tone in… | 26 | 4204 | maintenance |
| martinpacesa/BindCraft BindCraft is a Python-based computational pipeline for de novo protein binder design that combines AlphaFold2 backpropagation, ProteinMPNN,… | 77 | 1201 | active |
| nishuzumi/gemini-teacher A Python CLI application that acts as an English speaking practice assistant powered by Google Gemini. It listens to your speech via microp… | 68 | 1200 | active |