domain: artificial-intelligence
4539 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| lucidrains/tab-transformer-pytorch A PyTorch implementation of the TabTransformer architecture, an attention-based neural network for tabular data, also including the FT Tran… | 69 | 1092 | stable |
| Bogdanovich77/DeekSeek-OCR---Dockerized-API A Dockerized REST API and batch processing scripts that convert PDF documents to Markdown using the DeepSeek-OCR model behind a FastAPI bac… | 38 | 1092 | active |
| HacxGPT-Official/HacxGPT-CLI HacxGPT-CLI is a Python command-line chat client for accessing uncensored, unrestricted language models, including HacxGPT's own paid API m… | 79 | 1091 | active |
| hinthornw/trustcall A Python library that improves LLM tool calling reliability by having models generate JSON patch operations instead of full JSON blobs. It … | 38 | 1091 | active |
| yerfor/Real3DPortrait Official PyTorch implementation of Real3D-Portrait, an ICLR 2024 Spotlight paper for one-shot realistic 3D talking portrait synthesis. It g… | 26 | 1091 | active |
| wenge-research/x-agent X-Agent (智川) is an open-source, self-hosted enterprise agent development platform by Wenge that lets users build AI applications with zero … | 34 | 1088 | active |
| lastmile-ai/aiconfig AIConfig is an open-source framework that manages generative AI prompts, models, and model parameters as JSON-serializable configs, separat… | 48 | 1087 | active |
| piotrostr/listen Listen is a Rust toolkit and framework for algorithmic trading on Solana, evolving into a framework for AI-driven crosschain portfolio mana… | 46 | 1085 | active |
| Alpha-VLLM/Lumina-mGPT-2.0 Lumina-mGPT 2.0 is a stand-alone decoder-only autoregressive model trained from scratch that unifies a broad range of image generation task… | 42 | 1085 | active |
| mlc-ai/web-stable-diffusion A project that compiles and runs Stable Diffusion text-to-image models entirely inside web browsers using WebGPU and WebAssembly, with no s… | 30 | 3723 | maintenance |
| NVlabs/Fast-dLLM NVIDIA's official implementation of Fast-dLLM, a family of training-free and fine-tuning-based acceleration techniques for diffusion-based … | 57 | 1084 | active |
| StreetLamb/tribe Tribe AI is a low-code web application for rapidly building and coordinating multi-agent teams using drag-and-drop, built on LangChain and … | 48 | 1084 | active |
| olivia-ai/olivia Olivia is an open-source chatbot written in Go that uses a neural network for natural language understanding, aiming to be a free alternati… | 10 | 3717 | maintenance |
| spring-ai-alibaba/Lynxe Lynxe (formerly JManus) is a Java-based, code-free 'Prompt Programming' studio implementing a Manus-style multi-agent system with high exec… | 76 | 1081 | active |
| GradientHQ/symphony-coord Symphony-Coord is a decentralized multi-agent framework that treats agent selection as an online multi-armed bandit problem, allowing roles… | 54 | 1081 | active |
| XInTheDark/raycast-g4f A Raycast extension that provides free access to GPT-4, Claude, Llama, Gemini and other AI models via GPT4Free providers, with full support… | 56 | 1080 | active |
| minghanqin/LangSplat Official implementation of LangSplat, a CVPR 2024 Highlight paper that constructs a 3D language field using 3D Gaussian Splatting with CLIP… | 47 | 1080 | active |
| jae-jae/fetcher-mcp A Model Context Protocol (MCP) server that fetches web page content using a Playwright headless browser, executing JavaScript to handle dyn… | 47 | 1080 | active |
| auduno/headtrackr headtrackr is a JavaScript library for real-time face tracking and head tracking via a webcam using WebRTC/getUserMedia. It estimates the u… | 32 | 3700 | maintenance |
| XiaomiMiMo/MiMo-Audio Xiaomi's open-source 7B audio language model family (Base and Instruct) plus a 1.2B RVQ audio tokenizer, pretrained on 100M+ hours of audio… | 56 | 1077 | active |
| Pokee-AI/PokeeResearchOSS An open-source repository for Pokee's 7B-parameter DeepResearch agent that performs multi-turn web search and content reading to answer com… | 38 | 1077 | active |
| titanwings/ex-skill A Python tool that generates Claude Code / OpenClaw agent 'skills' capturing a specific person's texting persona from chat history (WeChat,… | 61 | 1076 | active |
| trailofbits/anamorpher Anamorpher is a tool for crafting and visualizing image scaling attacks that hide multi-modal prompt injections in images, revealed only wh… | 55 | 1076 | active |
| jjsantos01/qgis_mcp A Model Context Protocol integration that lets LLMs like Claude control QGIS Desktop via a socket-based QGIS plugin and an MCP server. It s… | 39 | 1076 | active |
| mit-han-lab/streaming-vlm StreamingVLM is a vision-language model framework from MIT Han Lab for real-time understanding of effectively infinite video streams. It ma… | 38 | 1076 | active |
| openai/glide-text2im Official codebase for GLIDE, a diffusion-based text-conditional image synthesis model from OpenAI. It provides pretrained models and notebo… | 10 | 3685 | maintenance |
| orailnoor/cross-platform-llm-client PrivateLM is a cross-platform AI chat client built with Flutter that unifies local on-device LLM inference (GGUF models with Vulkan GPU acc… | 72 | 1074 | active |
| mskayyali/nodepad nodepad is a spatial note-taking web application where notes are placed on a canvas and AI quietly classifies them, infers connections, and… | 56 | 1074 | active |
| Applio Applio is an open-source, MIT-licensed voice conversion suite built on RVC that lets users convert audio into other voices, train custom vo… | 90 | 3677 | maintenance |
| aklofas/kicad-happy A collection of AI coding agent skills that turn Claude Code, OpenAI Codex, and similar agents into KiCad electronics design assistants. It… | 82 | 1073 | active |
| OranAi-Ltd/oransim Oransim is an open-source causal digital twin engine for marketing, simulating a large virtual consumer society with LLM-backed personas to… | 56 | 1073 | active |
| AILab-CVC/UniRepLKNet UniRepLKNet is a large-kernel ConvNet architecture (CVPR 2024, TPAMI 2025) that provides universal perception across image, audio, video, p… | 42 | 1073 | stable |
| memoavatar/memo MEMO is an open-weight diffusion model for generating expressive, identity-consistent talking videos from a single reference image and an a… | 40 | 1072 | active |
| GanymedeNil/document.ai A universal local knowledge base solution that stores documents as vectors in a vector database and uses GPT3.5 to generate answers from re… | 30 | 3669 | maintenance |
| tryAGI/LangChain LangChain .NET is a C#/.NET implementation of the LangChain framework for building LLM-powered applications through composability. It provi… | 64 | 1071 | active |
| sipeed/TinyMaix TinyMaix is a tiny neural network inference library for microcontrollers (TinyML), with core code under 400 lines and a .text section under… | 34 | 1071 | active |
| trendy-design/llmchat LLMChat is a privacy-focused AI chat platform built with Next.js and TypeScript that provides a unified interface for chatting with multipl… | 24 | 1071 | active |
| juntang-zhuang/Adabelief-Optimizer AdaBelief is a deep learning optimizer that adapts step sizes based on the 'belief' in observed gradients, combining Adam's fast convergenc… | 32 | 1070 | stable |
| Stability-AI/stable-point-aware-3d SPAR3D is Stability AI's open-source model for fast single-image 3D mesh reconstruction using a two-stage pipeline with point cloud conditi… | 30 | 1070 | active |
| brenpoly/be-more-agent An offline-first conversational AI agent that turns a Raspberry Pi into a fully local voice assistant. It combines OpenWakeWord, Whisper.cp… | 64 | 1069 | active |
| codeacme17/examor Examor is a self-hosted web application that generates exams and review questions from your knowledge notes using LLMs like GPT-4 and Claud… | 31 | 1069 | active |
| ototadana/sd-face-editor A Stable Diffusion Web UI extension that detects and regenerates faces in generated images to fix broken faces, change facial expressions, … | 23 | 1069 | active |
| dynamiq-ai/dynamiq Dynamiq is an open-source Python orchestration framework for building agentic AI and LLM applications, covering agents, RAG pipelines, work… | 85 | 1066 | active |
| open-webui/openapi-servers A collection of reference OpenAPI Tool Server implementations that expose external tools and data sources to LLM agents over standard REST … | 39 | 1066 | active |
| metaskills/experts Experts.js is a JavaScript library that simplifies creating and deploying OpenAI Assistants via the Assistants API, hiding Run object manag… | 25 | 1064 | active |
| Lazarus-AI/clearwing Clearwing is a dual-mode autonomous offensive-security tool that combines a network-pentest ReAct agent with an LLM-driven source-code vuln… | 63 | 1063 | active |
| pepperoni21/ollama-rs A Rust client library for interacting with the Ollama API, supporting completion generation, streaming, chat, model management, embeddings,… | 88 | 1062 | active |
| AIrjen/OneButtonPrompt OneButtonPrompt is a Python tool that generates complete Stable Diffusion prompts from scratch with a single click, for beginners and users… | 44 | 1062 | active |
| vijishmadhavan/ArtLine ArtLine is a deep learning project that converts portrait photos into line art portraits, with a ControlNet-based variant that adjusts styl… | 32 | 3631 | maintenance |
| Sollimann/bonsai Bonsai is a Rust implementation of behavior trees for building deterministic AI logic, with Python bindings available on PyPI. It targets r… | 89 | 1061 | active |
| zai-org/GLM-TTS GLM-TTS is a Python-based text-to-speech synthesis system built on large language models, using a two-stage LLM plus Flow model architectur… | 50 | 1061 | active |
| JudgmentLabs/judgeval Judgeval is an open-source Python SDK for LLM agent improvement, providing OpenTelemetry-based tracing and prompt-based agent-judge evaluat… | 85 | 1060 | active |
| AnnaSuSu/TechSpar TechSpar is an open-source AI-powered technical interview preparation application that combines targeted training, resume-based mock interv… | 82 | 1060 | active |
| yokingma/SearChat SearChat is a self-hostable AI-powered conversational search engine that combines web search (Bing, Google, SearXNG) with multi-turn chat b… | 87 | 1059 | active |
| OpenLAIR/dr-claw Dr. Claw is an open-source, model-agnostic AI research assistant that provides a desktop and mobile UI for Claude Code, Cursor CLI, Codex, … | 82 | 1059 | active |
| RUC-NLPIR/Arbor Arbor is a generalist autonomous research agent that takes a benchmark and a goal, then proposes hypotheses, edits code, and runs real expe… | 79 | 1059 | active |
| danijar/dreamerv2 A TensorFlow 2 implementation of the DreamerV2 model-based reinforcement learning agent that learns world models from high-dimensional imag… | 32 | 1058 | stable |
| AITuberKit AITuberKit is an all-in-one web application toolkit for building and deploying AI character chat experiences, including streaming-oriented … | 88 | 1056 | active |
| vstorm-co/pydantic-deepagents Pydantic Deep Agents is an open-source, self-hosted terminal AI assistant (a Claude Code alternative) plus the Python framework behind it, … | 77 | 1056 | active |
| aiptimizer/TurboOCR TurboOCR is an extremely fast GPU-accelerated document parser written in C++ that combines OCR, layout analysis, table extraction, and form… | 82 | 1055 | active |
| shang-zhu/violin Violin is an open-source video translation tool that transcribes speech, translates it into 33 languages, synthesizes a native-sounding voi… | 59 | 1055 | active |
| Tencent-Hunyuan/HunyuanVideo-Foley HunyuanVideo-Foley is a multimodal diffusion model from Tencent Hunyuan that generates high-fidelity Foley sound effects synchronized with … | 37 | 1055 | active |
| GaussianAvatars Official research code for GaussianAvatars, a CVPR 2024 Highlight method that creates photorealistic, fully controllable head avatars by ri… | 56 | 1054 | active |
| CRui5in/paper-ppt-agent A self-hosted web application that uses a multi-agent AI pipeline (Strategist, Executor, Critic) to convert academic papers from PDF or LaT… | 78 | 1053 | active |
| InternRobotics/PointLLM PointLLM is a multimodal large language model that understands colored 3D point clouds of objects, built on a point cloud encoder fused wit… | 64 | 1053 | active |
| showlab/MotionDirector MotionDirector is a research library for customizing text-to-video diffusion models to generate videos with desired motions from a small se… | 27 | 1053 | active |
| BAAI-DCAI/Bunny Bunny is a family of lightweight multimodal vision-language models that combine plug-and-play vision encoders (EVA-CLIP, SigLIP) with langu… | 26 | 1053 | active |
| muapi CLI muapi-cli and its Generative Media Skills provide a schema-driven CLI, skill library, and MCP server that let AI agents (Claude Code, Curso… | 93 | 1051 | active |
| Artrajz/vits-simple-api A Python HTTP API service that exposes VITS-family text-to-speech models (VITS, Bert-VITS2, GPT-SoVITS, W2V2 emotional VITS) for inference.… | 69 | 1051 | active |
| openai/retro Gym Retro is a Python library that turns classic video games into Gym environments for reinforcement learning research, with integrations f… | 10 | 3584 | maintenance |
| LAM LAM is a PyTorch implementation of a Large Avatar Model that reconstructs an animatable 3D Gaussian head from a single image in one forward… | 58 | 1050 | active |
| nv-tlabs/PiD PiD is a plug-and-play pixel diffusion decoder from NVIDIA that replaces VAE/RAE decoders, decoding latent representations directly into hi… | 55 | 1050 | active |
| shyamsaktawat/OpenAlpha_Evolve OpenAlpha_Evolve is an open-source Python framework inspired by DeepMind's AlphaEvolve that uses LLMs to iteratively generate, test, and ev… | 29 | 1050 | active |
| agents-flex/agents-flex Agents-Flex is a lightweight, modular Java framework for building AI applications and agents, positioned as a Java counterpart to Spring AI… | 93 | 1049 | active |
| chAng-L19/codex-redteam-mode An opt-in red-team mode plugin for OpenAI Codex App and Codex CLI that compiles offensive-security objectives into GoalContracts and execut… | 79 | 1049 | active |
| HannesStark/boltzgen BoltzGen is an open-source all-atom generative diffusion model for designing protein and peptide binders against arbitrary biomolecular tar… | 70 | 1049 | active |
| zai-org/SCAIL SCAIL is the official inference implementation of a 14B diffusion transformer model that generates studio-grade character animation videos … | 52 | 1048 | active |
| wu-yc/LabClaw LabClaw is a modular library of 240 SKILL.md agent skills for biomedical AI research, covering biology, drug discovery, medicine, data scie… | 47 | 1048 | active |
| PatterAI/Patter Patter is an open-source, MIT-licensed SDK (Python and TypeScript) that connects AI agents to real phone calls, handling telephony, speech-… | 78 | 1047 | active |
| kernel/kernel-images Kernel is an open-source browser infrastructure platform that provides sandboxed, ready-to-use Chromium browsers as a service for browser a… | 65 | 1047 | active |
| alexfrom0815/Online-3D-BPP-PCT A Python research implementation of the ICLR 2022 paper 'Learning Efficient Online 3D Bin Packing on Packing Configuration Trees', using de… | 60 | 1047 | active |
| octos-org/octos Octos is a self-hosted AI agent runtime that runs a personal AI assistant on your own machine, connecting to providers like Anthropic, Open… | 80 | 1046 | active |
| airbnb/chronon Chronon is an open-source data platform from Airbnb for computing, backfilling, and serving ML features. It handles batch and streaming fea… | 75 | 1046 | active |
| kreneskyp/ix IX is a self-hosted platform for designing, running, and scaling autonomous and semi-autonomous LLM-powered agents and workflows. It includ… | 47 | 1046 | active |
| Minima-AI-Inc/minima Minima is an open-source, on-premises conversational RAG system deployed via Docker containers. It supports fully local operation with Olla… | 40 | 1046 | active |
| 3DTopia/3DTopia-XL 3DTopia-XL is a 3D diffusion transformer model that generates high-quality 3D assets with PBR materials from a single image or text prompt … | 37 | 1046 | active |
| libtv-labs/libtv-skills A collection of AI agent skill packages that expose LibLib.tv's AIGC capabilities (AI image and video generation) via its OpenAPI. It follo… | 47 | 1045 | active |
| rayenfeng/riko_project Project Riko is an anime-themed conversational voice assistant that combines OpenAI's GPT for dialogue, GPT-SoVITS for voice synthesis, and… | 31 | 1045 | active |
| Nathan-code-development/AIApplication A cross-platform mobile chat application built with .NET MAUI in C# that lets users converse with multiple AI models including DeepSeek, Do… | 60 | 1044 | active |
| splx-ai/agentic-radar Agentic Radar is an open-source security scanner that analyzes LLM agentic workflows built with popular frameworks (e.g., CrewAI, LangGraph… | 52 | 1044 | active |
| Tencent-Hunyuan/InstantCharacter InstantCharacter is a tuning-free framework built on diffusion transformers that generates character-consistent images from a single refere… | 28 | 1044 | active |
| star23/Day1Global-Skills A collection of investment analysis skills for AI agents, covering tech earnings deep dives, Buffett-style value investing, US market senti… | 58 | 1043 | active |
| metauto-ai/GPTSwarm GPTSwarm is a Python library for building LLM-based agents as computational graphs, with modules for agent graphs, memory, LLM backends, an… | 55 | 1042 | active |
| Nutlope/blinkshot BlinkShot is an open-source web application that generates AI images in real time as you type, powered by the Flux Schnell model via Togeth… | 67 | 1040 | active |
| darrencxl0301/StageRAG StageRAG is a Python framework/blueprint for building production-ready RAG applications with switchable 3-step (Speed) and 4-step (Precisio… | 38 | 1040 | active |
| tensorflow/minigo Minigo is an open-source, minimalist implementation of the AlphaGo Zero algorithm for the game of Go, built on TensorFlow. It provides a re… | 10 | 3542 | maintenance |
| RLCard RLCard is a Python toolkit for reinforcement learning research in card games, providing environments for Blackjack, Leduc Hold'em, Texas Ho… | 23 | 3541 | maintenance |
| kijai/ComfyUI-Hunyuan3DWrapper A ComfyUI custom node wrapper for Tencent's Hunyuan3D-2 model, enabling 3D asset generation from images or text directly inside ComfyUI wor… | 53 | 1039 | active |
| D2I-CUHKSZ/MicroWorld MicroWorld is a lightweight Python engine that turns multi-modal event materials (documents, images, videos, graph signals) into structured… | 52 | 1039 | active |
| ricklamers/gpt-code-ui An open-source, self-hosted implementation of OpenAI's ChatGPT Code Interpreter. Users chat with GPT-3.5/GPT-4 models that generate and exe… | 20 | 3537 | maintenance |