domain: artificial-intelligence
4539 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| red42/HTML5_Genetic_Cars A browser-based simulation that evolves 2D cars using a genetic algorithm rendered on HTML5 canvas with the Box2D physics engine. It runs e… | 10 | 1226 | maintenance |
| ml4a/ml4a-ofx A collection of openFrameworks applications in C++ for real-time interactive machine learning, aimed at artists and creative coders. It inc… | 23 | 1223 | maintenance |
| iqiyi/FASPell FASPell is a Python-based Chinese spell checker based on the DAE-Decoder paradigm, published at the EMNLP 2019 W-NUT workshop. It detects a… | 32 | 1222 | maintenance |
| GeekAlexis/FastMOT FastMOT is a high-performance multiple object tracking system combining YOLO/SSD detection, Deep SORT with OSNet ReID, and KLT optical flow… | 23 | 1220 | maintenance |
| timqian/openprompt.co OpenPrompt.co is a web application for creating, using, and sharing ChatGPT prompts, with a community-driven catalog of starred prompts. Th… | 74 | 1218 | maintenance |
| TensorFlow-Unreal An Unreal Engine plugin that wraps TensorFlow machine learning operations as Actor Components accessible from C++, Blueprints, and Python. … | 23 | 1218 | maintenance |
| ucbdrive/few-shot-object-detection FsDet is the official implementation of the ICML 2020 paper 'Frustratingly Simple Few-Shot Object Detection' (TFA), built on detectron2. It… | 23 | 1218 | maintenance |
| UMass-Embodied-AGI/3D-LLM 3D-LLM is the research code for a large language model that takes 3D representations (objects and scenes) as input, built on BLIP-2/LAVIS. … | 28 | 1213 | maintenance |
| facebookresearch/BLINK BLINK is a Python entity linking library from Facebook Research that resolves mentions in text to Wikipedia entities using a two-stage bi-e… | 10 | 1210 | maintenance |
| NVlabs/VoxFormer Official PyTorch implementation of VoxFormer, a CVPR 2023 highlight paper presenting a sparse voxel transformer for camera-based 3D semanti… | 31 | 1208 | maintenance |
| mozilla-ai/any-agent any-agent is a Python library providing a single unified interface for building, serving, and evaluating AI agents across multiple agent fr… | 72 | 1200 | maintenance |
| KwaiKEG/KwaiAgents KwaiAgents is an open-source suite from Kuaishou for building generalized information-seeking agents powered by LLMs. It includes KAgentSys… | 27 | 1200 | maintenance |
| facebookresearch/House3D House3D is a virtual 3D environment of over 45k fully annotated indoor scenes from the SUNCG dataset, built for training embodied AI agents… | 10 | 1200 | maintenance |
| DeepWisdom/AutoDL AutoDL is a fully automated deep learning framework from DeepWisdom that won the NeurIPS AutoDL challenge. It performs automatic feature en… | 23 | 1197 | maintenance |
| lucidrains/CoCa-pytorch A Pytorch implementation of CoCa (Contrastive Captioners), an image-text foundation model that combines contrastive learning with an encode… | 23 | 1196 | maintenance |
| rinongal/StyleGAN-nada Official PyTorch implementation of StyleGAN-NADA, a CLIP-guided method for adapting pre-trained StyleGAN image generators to new domains us… | 32 | 1195 | maintenance |
| LiuHC0428/LAW-GPT A Chinese-language legal conversational model (LawGPT_zh / 獬豸) fine-tuned from ChatGLM-6B with LoRA on legal Q&A datasets, including answer… | 30 | 1195 | maintenance |
| altosaar/variational-autoencoder A reference implementation of a variational autoencoder (VAE) in PyTorch, TensorFlow, and JAX, including an inverse autoregressive flow var… | 23 | 1192 | maintenance |
| visual-openllm/visual-openllm An open-source tool that interactively connects different visual models with an LLM, built on ChatGLM, Visual ChatGPT, and Stable Diffusion… | 30 | 1186 | maintenance |
| business-science/ai-data-science-team A Python library of specialized LLM-powered agents for common data science workflows such as data loading, cleaning, wrangling, visualizati… | 59 | 5394 | experimental |
| HyperGAN/HyperGAN HyperGAN is a composable GAN (generative adversarial network) framework built on PyTorch, offering both a Python API and a CLI with a user … | 23 | 1183 | maintenance |
| Mentra-Community/OpenSourceSmartGlasses An open source smart glasses hardware and software project with display, microphones, wireless phone connection, and prescription lenses, i… | 23 | 1182 | maintenance |
| lucidrains/performer-pytorch A PyTorch implementation of the Performer transformer, which uses linear attention via the FAVOR+ (Fast Attention Via positive Orthogonal R… | 23 | 1182 | maintenance |
| steven-tey/chathn ChatHN is an open-source AI chatbot web app that lets users query Hacker News using natural language, powered by OpenAI function calling an… | 29 | 1178 | maintenance |
| cubiq/ComfyUI_essentials A collection of essential custom nodes for ComfyUI that add new image processing and utility features missing from the core. The repository… | 34 | 1173 | maintenance |
| MLNLP-World/AI-Paper-Collector A tool from the MLNLP community that automatically collects AI-related papers from sources like DBLP, ACL Anthology, NIPS, and OpenReview, … | 32 | 1173 | maintenance |
| abhagsain/ai-cli A GPT-3/ChatGPT-powered command-line tool that answers CLI command questions directly from the terminal. Built with oclif in TypeScript and… | 23 | 1171 | maintenance |
| google/active-learning A Python module for running experiments comparing different active learning algorithms on benchmark datasets. It provides a main experiment… | 10 | 1163 | maintenance |
| ChenyangQiQi/FateZero FateZero is a zero-shot text-based video editing framework built on pretrained Stable Diffusion models, introduced in an ICCV 2023 Oral pap… | 21 | 1162 | maintenance |
| soulteary/docker-prompt-generator A Docker-based web application that uses language models to generate and expand prompts for image generation tools like MidJourney and Stab… | 30 | 1157 | maintenance |
| ajay-sainy/Wav2Lip-GFPGAN A pipeline combining Wav2Lip lip-sync generation with GFPGAN face restoration to produce high-quality talking-head videos from an input vid… | 32 | 1156 | maintenance |
| all-in-aigc/sorafm Sora.FM is a self-hostable Next.js web application that showcases AI-generated videos from OpenAI's Sora text-to-video model and provides a… | 25 | 1153 | maintenance |
| isekaidev/stable.art Stable.art is an open-source Photoshop plugin (v23.3.0+) that integrates Stable Diffusion image generation via an Automatic1111 API backend… | 22 | 1152 | maintenance |
| HackerPoet/Composer Composer is a neural network-based application that generates video game music. It is written in Python and produces chiptune-style game so… | 32 | 1151 | maintenance |
| showlab/Show-1 Show-1 is a research codebase implementing a text-to-video generation model that combines pixel and latent diffusion models, published at I… | 45 | 1148 | maintenance |
| pix2pixzero/pix2pix-zero pix2pix-zero is a Python library implementing zero-shot image-to-image translation using pre-trained Stable Diffusion models. It enables ed… | 31 | 1146 | maintenance |
| sentient-agi/ROMA ROMA (Recursive Open Meta-Agent) is a Python meta-agent framework for building hierarchical, high-performance multi-agent systems. It enabl… | 51 | 5176 | experimental |
| Shiriluz/Word-As-Image Official implementation of the Word-As-Image semantic typography technique (SIGGRAPH 2023), which automatically illustrates letters so they… | 30 | 1144 | maintenance |
| KnockOutEZ/wigolo wigolo is a local-first MCP server that gives AI coding agents web intelligence tools: search across 18 engines, fetch with anti-bot escala… | 80 | 5171 | experimental |
| lupantech/chameleon-llm Chameleon is a research framework for plug-and-play compositional reasoning with large language models, where an LLM planner synthesizes pr… | 20 | 1141 | maintenance |
| Open Interpreter 01 An open-source voice interface platform that lets users control computers conversationally, powered by Open Interpreter. It pairs a Python … | 17 | 5153 | experimental |
| google-research/mixmatch Reference implementation of MixMatch, a holistic semi-supervised learning algorithm from a Google Research paper. It includes dataset prepa… | 10 | 1139 | maintenance |
| IBM/Dromedary Dromedary is an open-source self-aligned language model trained with minimal human supervision using the principle-driven SELF-ALIGN pipeli… | 48 | 1137 | maintenance |
| fixie-ai/ai-jsx AI.JSX is a JavaScript/TypeScript framework for building AI applications using JSX components, supporting prompt engineering, tools, docume… | 29 | 1132 | maintenance |
| MRzzm/DINet DINet is the official PyTorch implementation of an AAAI 2023 paper on realistic face visually dubbing, which deforms and inpaints mouth reg… | 32 | 1128 | maintenance |
| yu-takagi/StableDiffusionReconstruction Research codebase reproducing Takagi and Nishimoto's CVPR 2023 method for reconstructing images a person viewed from fMRI brain activity us… | 30 | 1127 | maintenance |
| Harmonai-org/sample-generator A set of tools and Jupyter notebooks for training generative diffusion models on arbitrary audio samples, built around Dance Diffusion. It … | 32 | 1118 | maintenance |
| google-deepmind/dramatron Dramatron is a research tool from DeepMind that uses pre-trained large language models to hierarchically co-write theatre scripts and scree… | 32 | 1114 | maintenance |
| primaryobjects/AI-Programmer AI-Programmer is a C# experiment that uses genetic algorithms to automatically generate programs in a Turing-complete esoteric language (Br… | 47 | 1102 | maintenance |
| nicrusso7/rex-gym Rex-gym provides OpenAI Gym environments for the open-source SpotMicro quadruped robot, built on PyBullet simulation, along with a PPO lear… | 32 | 1102 | maintenance |
| Flode-Labs/vid2densepose A Python tool that applies the DensePose model to videos, producing color-coded part-index visualizations for each frame. Its output is des… | 26 | 1102 | maintenance |
| auspicious3000/autovc AUTOVC is a PyTorch implementation of a many-to-many non-parallel voice conversion framework that performs zero-shot voice style transfer u… | 32 | 1100 | maintenance |
| hongfz16/AvatarCLIP Official PyTorch implementation of AvatarCLIP, a SIGGRAPH 2022 research framework that generates and animates 3D human avatars from natural… | 32 | 1100 | maintenance |
| Syan-Lin/CyberWaifu CyberWaifu is a Python chatbot application that combines LLMs (ChatGPT, Claude) with TTS (edge-tts, Azure) to create realistic conversation… | 30 | 1100 | maintenance |
| gdquest-demos/godot-2d-space-game Harvester is a free and open-source top-down 2D space mining game built with the Godot game engine and GDQuest's steering AI framework. It … | 59 | 1098 | maintenance |
| jeeliz/jeelizWeboji A JavaScript/WebGL library for real-time face tracking and facial expression detection in the browser, using a neural network to detect 11 … | 32 | 1097 | maintenance |
| MiguelsPizza/WebMCP WebMCP (MCP-B) is a protocol and TypeScript SDK that exposes JavaScript functions in web pages to LLMs as MCP tools, via a Chrome extension… | 41 | 1094 | maintenance |
| maelfabien/Multimodal-Emotion-Recognition A real-time multimodal emotion recognition web app built with Flask that analyzes emotions from text, audio, and video inputs using deep le… | 32 | 1089 | maintenance |
| lukasHoel/text2room Text2Room is a research codebase that generates room-scale textured 3D meshes from a text prompt by leveraging pre-trained 2D text-to-image… | 30 | 1089 | maintenance |
| exorde-labs/exorde-client The Exorde client is a Python CLI worker node for the Exorde Network, a decentralized protocol where participants scrape social media and w… | 60 | 1084 | maintenance |
| agentcoinorg/evo.ninja Evo.ninja is a versatile generalist AI agent that adapts in real time by selecting pre-defined agent personas (researcher, CSV analyst, dev… | 28 | 1081 | maintenance |
| YudongGuo/AD-NeRF A PyTorch implementation of AD-NeRF, an ICCV 2021 paper that synthesizes talking-head videos by driving neural radiance fields with audio i… | 32 | 1072 | maintenance |
| Wangt-CN/DisCo DisCo is a CVPR 2024 research codebase for referring human dance generation, producing realistic dance images and videos from a reference h… | 29 | 1072 | maintenance |
| databricks/lilac Lilac is an open-source tool for exploring, curating, and quality-controlling datasets used for training, fine-tuning, and monitoring LLMs.… | 10 | 1072 | maintenance |
| thunderbird/thunderbolt Thunderbolt is an open-source, cross-platform AI client from Thunderbird that lets users connect any OpenAI-compatible model or ACP-compati… | 82 | 4760 | experimental |
| MatthieuCourbariaux/BinaryNet BinaryNet is a research codebase for training deep neural networks whose weights and activations are constrained to +1 or -1, reproducing t… | 32 | 1068 | maintenance |
| Stable Diffusion NCNN A C++ implementation of Stable Diffusion using the NCNN inference framework, supporting both txt2img and img2img. It runs on x86 Windows ex… | 23 | 1067 | maintenance |
| ntasfi/PyGame-Learning-Environment PyGame Learning Environment (PLE) is a Python library providing a reinforcement learning environment with a suite of PyGame-based games, mi… | 32 | 1066 | maintenance |
| omerbt/MultiDiffusion Official PyTorch implementation of MultiDiffusion (ICML 2023), a training-free framework that fuses multiple diffusion paths over a pre-tra… | 31 | 1066 | maintenance |
| JiehangXie/PaddleBoBo PaddleBoBo is a Python project built on PaddlePaddle (with PaddleSpeech and PaddleGAN) that quickly generates a virtual streamer (VTuber) f… | 32 | 1063 | maintenance |
| OpenBMB/VisCPM VisCPM is a family of open-source bilingual (Chinese/English) multimodal large models built on the 10B CPM-Bee language model, comprising V… | 29 | 1062 | maintenance |
| jsyoon0823/TimeGAN Reference implementation of TimeGAN, a generative adversarial network framework for generating synthetic time-series data, published at Neu… | 62 | 1060 | maintenance |
| OFA-Sys/ONE-PEACE ONE-PEACE is a general multimodal representation model that jointly encodes vision, audio, and language modalities without initializing fro… | 29 | 1060 | maintenance |
| featurecat/lizzie Lizzie is a Java-based graphical interface for analyzing Go (baduk) games in real time using the Leela Zero engine. It displays win rates, … | 23 | 1060 | maintenance |
| obiscr/ChatGPT A JetBrains IDE plugin that integrates ChatGPT into IntelliJ-based IDEs, providing an in-editor chat interface with OpenAI's language model… | 32 | 1059 | maintenance |
| SeetaFace SeetaFace is an open-source, full-stack face recognition toolkit written in standard C++ with no third-party dependencies. It provides face… | 32 | 1058 | maintenance |
| raminmh/CfC Reference implementations of Closed-form Continuous-time (CfC) neural networks, a fast closed-form approximation of liquid time-constant ne… | 23 | 1054 | maintenance |
| Edresson/YourTTS YourTTS is a zero-shot multi-speaker text-to-speech and voice conversion model built on VITS, implemented in the Coqui TTS framework. It su… | 23 | 1053 | maintenance |
| microsoft/Oscar Oscar is Microsoft's research code for object-semantics aligned cross-modal pre-training of vision-language models, with VinVL providing im… | 10 | 1053 | maintenance |
| mpaepper/llm_agents A small Python library for building agents controlled by large language models, inspired by LangChain but implemented from scratch in very … | 42 | 1052 | maintenance |
| ashkamath/mdetr MDETR (Modulated Detection) is a PyTorch research codebase for end-to-end multi-modal object detection that grounds free-form text queries … | 32 | 1052 | maintenance |
| microsoft/Cognitive-Samples-IntelligentKiosk A UWP sample application from Microsoft showcasing hands-free kiosk-style demos built on Azure Cognitive Services (Face, Computer Vision, T… | 10 | 1052 | maintenance |
| jondurbin/airoboros Airoboros is a Python library implementing a heavily modified version of the Self-Instruct paper to generate high-quality synthetic instruc… | 20 | 1051 | maintenance |
| bravekingzhang/text2video A Python web application that converts text (e.g., novel passages) into narrated videos. It splits text into sentences, generates images vi… | 29 | 1049 | maintenance |
| CSHaitao/LexiLaw LexiLaw is a fine-tuned Chinese legal large language model based on ChatGLM-6B, providing legal consultation and Q&A capabilities. It inclu… | 61 | 1043 | maintenance |
| pixray/pixray Pixray is a Python library and command-line utility for text-to-image generation, combining CLIP-guided GAN imagery, pixel-art drawers, and… | 23 | 1042 | maintenance |
| thuml/Anomaly-Transformer Official PyTorch implementation of the Anomaly Transformer model from the ICLR 2022 Spotlight paper on unsupervised time series anomaly det… | 32 | 1041 | maintenance |
| Xwin-LM/Xwin-LM Xwin-LM is an open-source project for LLM alignment technologies including supervised fine-tuning, reward models, reject sampling, and RLHF… | 28 | 1036 | maintenance |
| wywu/LAB Official C++/Caffe implementation of the CVPR 2018 paper 'Look at Boundary: A Boundary-Aware Face Alignment Algorithm', which localizes fac… | 32 | 1018 | maintenance |
| CaviraOSS/OpenMemory OpenMemory is a local-first, self-hosted cognitive memory engine that gives LLM applications and AI agents persistent long-term memory, wit… | 70 | 4482 | experimental |
| trishume/eyeLike eyeLike is an OpenCV-based C++ implementation of Fabian Timm's gradient-based eye center localization algorithm for webcam pupil tracking. … | 32 | 1015 | maintenance |
| ggeop/Python-ai-assistant Jarvis is a Python voice-controlled AI assistant for Linux that recognizes speech, responds conversationally, and executes commands like op… | 23 | 1014 | maintenance |
| dimensionalOS/dimos DimOS is a Python SDK and agent-native operating system for generalist robotics, letting users command humanoids, quadrupeds, and drones in… | 86 | 4460 | experimental |
| tensorflow/neural-structured-learning Neural Structured Learning (NSL) is a TensorFlow framework for training neural networks with structured signals, either explicit graphs or … | 64 | 1011 | maintenance |
| PAIR-code/what-if-tool The What-If Tool (WIT) is a visual interface from Google's PAIR team for probing black-box classification and regression ML models without … | 62 | 1011 | maintenance |
| varunshenoy/GraphGPT GraphGPT is a web application that converts unstructured natural language text into a knowledge graph using GPT-3, visualizing entities and… | 31 | 4427 | experimental |
| cbh123/narrator A Python app that watches your webcam and generates David Attenborough-style narration of what it sees, using GPT vision models and ElevenL… | 69 | 4426 | experimental |
| fcakyon/autollm AutoLLM is a Python library for rapidly building RAG-based LLM web apps and APIs, offering a unified API over 100+ LLM providers and 20+ ve… | 10 | 1005 | maintenance |
| aws-samples/bedrock-access-gateway A proxy service that exposes Amazon Bedrock foundation models through OpenAI-compatible RESTful APIs, letting existing OpenAI SDK-based app… | 10 | 1005 | maintenance |
| susiai/susi_device SUSI Device provides sources to install the SUSI AI assistant stack on a Raspberry Pi, combining microphone/speaker, a small display, a loc… | 32 | 1004 | maintenance |