domain: artificial-intelligence
4539 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| facebookresearch/fast3r Fast3R is the official PyTorch implementation of a CVPR 2025 model from Meta FAIR that reconstructs 3D scenes and estimates camera poses fr… | 10 | 1591 | active |
| Korea Investment Open Trading API Official sample code repository from Korea Investment & Securities for their Open Trading API (KIS Developers), providing Python examples f… | 76 | 1590 | active |
| Drexubery/ViewCrafter ViewCrafter is a research codebase that uses video diffusion models to synthesize high-fidelity novel views of scenes from a single or spar… | 48 | 1590 | active |
| meta-pytorch/torchtune Torchtune is a PyTorch-native library for authoring, post-training, and experimenting with large language models. It provides hackable trai… | 70 | 5802 | maintenance |
| gcorso/DiffDock DiffDock is a deep learning implementation of a diffusion generative model for molecular docking, predicting how small molecule ligands bin… | 31 | 1570 | active |
| roryclear/clearcam Clearcam is a self-hosted Python NVR that adds AI object detection, tracking, mobile notifications, and semantic search to any RTSP securit… | 84 | 1568 | active |
| openilink/openilink-hub OpeniLink Hub is a self-hosted, open-source message management platform and app marketplace for WeChat ClawBot (iLink protocol) bots, built… | 72 | 1565 | active |
| feenkcom/gtoolkit Glamorous Toolkit is the Moldable Development environment built on Pharo/Smalltalk, providing an integrated IDE with thousands of contextua… | 96 | 1562 | active |
| Roncoo Education System Roncoo Education (领课教育系统) is an open-source online education platform built with a Spring Cloud Alibaba microservices backend and Vue 3/Nux… | 68 | 1546 | active |
| RLHFlow/RLHF-Reward-Modeling A collection of training recipes for reward models used in RLHF, covering Bradley-Terry reward models, pairwise preference models, ArmoRM, … | 33 | 1541 | active |
| ROMP ROMP is a PyTorch-based library and pip-installable API (simple-romp) for real-time monocular multi-person 3D human mesh recovery, implemen… | 23 | 1539 | stable |
| feremabraz/bloomberg-terminal A Bloomberg Terminal clone built with Next.js 15, React 19, and TypeScript that provides real-time financial market data visualization in a… | 50 | 1534 | active |
| BrokenSource/DepthFlow DepthFlow is a free, open-source Python application and library that converts still images into 3D parallax effect videos using monocular d… | 84 | 1528 | active |
| Tencent/TFace TFace is a research platform from Tencent Youtu Lab for trusty face analysis, covering face recognition, face security (anti-spoofing), fac… | 57 | 1523 | active |
| allenzren/open-pi-zero An open-source re-implementation of the pi0 vision-language-action (VLA) model from Physical Intelligence, built on a pre-trained PaliGemma… | 24 | 1523 | active |
| decoderesearch/SAELens SAELens is a Python library for training sparse autoencoders (SAEs) on language model activations and analyzing them for mechanistic interp… | 89 | 1521 | active |
| PrathamLearnsToCode/paper2code An agent skill for coding agents like Claude Code that converts an arXiv paper URL into a working Python implementation. It generates citat… | 48 | 1517 | active |
| zapdos-labs/unblink Unblink is an AI-powered camera monitoring application that uses a vision language model (Qwen3-VL) to analyze camera frames, summarize act… | 50 | 1507 | active |
| edmund-io/edmunds-claude-code A Claude Code plugin providing 14 slash commands and 11 specialized AI agents for web development workflows. It scaffolds APIs, React compo… | 38 | 1502 | active |
| pluralsh/plural Plural is an enterprise Kubernetes management platform that provides fleet-scale GitOps deployments, infrastructure-as-code management, and… | 95 | 1497 | active |
| CUT3R/CUT3R CUT3R is the official PyTorch implementation of 'Continuous 3D Perception Model with Persistent State' (CVPR 2025 Oral), a stateful recurre… | 38 | 1489 | active |
| fawney19/Aether Aether is a self-hosted AI API gateway written in Rust that provides a unified entry point for Claude, OpenAI, Gemini, and their CLI client… | 81 | 1462 | active |
| LocoMuJoCo LocoMuJoCo is an imitation learning benchmark for whole-body locomotion control built on MuJoCo, featuring humanoid, quadruped, and musculo… | 75 | 1454 | active |
| zsyOAOA/InvSR InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff… | 50 | 1452 | active |
| tianweiy/DMD2 DMD2 is the official PyTorch implementation of Improved Distribution Matching Distillation, a NeurIPS 2024 method that distills diffusion m… | 28 | 1448 | active |
| writer/writer-framework Writer Framework is an open-source Python framework for building data and AI applications with a drag-and-drop visual editor for the fronte… | 76 | 1447 | active |
| aeon-toolkit/aeon aeon is a scikit-learn compatible Python toolkit for machine learning on time series, covering classification, regression, clustering, fore… | 87 | 1440 | active |
| neuralchen/SimSwap SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It… | 23 | 5186 | maintenance |
| honeyandme/RAGQnASystem A medical intelligent question-answering system combining knowledge-graph RAG with large language models, built on the DiseaseKG dataset wi… | 62 | 1437 | active |
| YanjieZe/3D-Diffusion-Policy 3D Diffusion Policy (DP3) is a visual imitation learning algorithm that combines compact 3D point cloud representations with diffusion poli… | 46 | 1437 | active |
| google/oss-fuzz-gen A Google framework that uses large language models to automatically generate fuzz targets for real-world C/C++, Java, and Python projects, … | 58 | 1434 | active |
| Francis-Rings/StableAnimator StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po… | 40 | 1431 | active |
| zsyOAOA/ResShift ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi… | 61 | 1427 | active |
| acon96/home-llm A Home Assistant custom integration plus fine-tuned small language models that let you control your smart home entirely with a local LLM, n… | 86 | 1426 | active |
| apple/ml-aim Apple's official repository for AIM (Autoregressive Image Models), providing code and pretrained checkpoints for AIMv1 and AIMv2 large visi… | 41 | 1424 | active |
| jakobhoeg/nextjs-ollama-llm-ui A fully-featured, ChatGPT-inspired web interface for chatting with local Ollama LLMs, built with Next.js, React, and Tailwind. It runs full… | 30 | 1423 | active |
| WPeGPT WPeGPT is an IDA Pro plugin that integrates LLM models (OpenAI, DeepSeek, or any OpenAI-compatible API) into binary analysis workflows. It … | 74 | 1422 | active |
| SagiPolaczek/NeuralSVG Official PyTorch implementation of NeuralSVG, an ICCV 2025 paper that generates layered, editable SVG vector graphics from text prompts. It… | 46 | 1419 | active |
| dexmal/dexbotic Dexbotic is an open-source PyTorch-based toolbox for developing Vision-Language-Action (VLA) models for embodied intelligence. It unifies p… | 74 | 1417 | active |
| nv-tlabs/GEN3C GEN3C is NVIDIA's research codebase for a generative video model that achieves precise camera control and temporal 3D consistency using a 3… | 59 | 1414 | active |
| lucidrains/self-rewarding-lm-pytorch A PyTorch library implementing the Self-Rewarding Language Model training framework from MetaAI, along with the SPIN training method. It pr… | 16 | 1410 | active |
| open-gigaai/giga-world-policy GigaWorld-Policy is a World Action Model (WAM) for robot policy learning that jointly models actions and future visual observations during … | 58 | 1403 | active |
| yfeng95/PRNet PRNet is a Python/TensorFlow implementation of the ECCV 2018 Position Map Regression Network for joint 3D face reconstruction and dense ali… | 32 | 5014 | maintenance |
| Junyi42/monst3r MonST3R is the official PyTorch implementation of an ICLR 2025 paper that estimates per-timestep geometry (pointmaps) from dynamic videos i… | 36 | 1387 | active |
| yanx27/Pointnet_Pointnet2_pytorch A pure PyTorch implementation of the PointNet and PointNet++ deep learning architectures for point cloud processing. It includes training a… | 32 | 4945 | maintenance |
| OpenPPL/ppl.nn PPLNN is a high-performance deep-learning inference engine written in C++ that runs ONNX models on x86 CPUs and NVIDIA GPUs, with a dedicat… | 32 | 1367 | active |
| andyhuo520/aetherviz-master AetherViz Master is an AI-powered interactive educational visualization tool that turns any teaching topic into an immersive 3D interactive… | 50 | 1365 | active |
| 79E/ChatGpt-Web A commercially-viable ChatGPT web application built with React and TypeScript, featuring a full admin backend for managing users, tokens, p… | 26 | 1363 | active |
| Sense-X/Co-DETR Co-DETR is a PyTorch implementation of DETRs with Collaborative Hybrid Assignments Training, an ICCV 2023 object detection and instance seg… | 32 | 1360 | stable |
| lxtGH/OMG-Seg Official research codebase for OMG-Seg (CVPR 2024) and OMG-LLaVA (NeurIPS 2024), unified models for image-level, object-level, and pixel-le… | 47 | 1354 | active |
| wyhuai/DDNM DDNM is a Python research codebase implementing the Denoising Diffusion Null-Space Model for zero-shot image restoration, published as an I… | 32 | 1351 | stable |
| emonney/QuickApp QuickApp is an opinionated full-stack project template combining Angular 21 and ASP.NET Core 10 with pre-built authentication, authorizatio… | 82 | 1350 | active |
| vercel-labs/gemini-chatbot An open-source Next.js chatbot template powered by Google Gemini and the Vercel AI SDK, with streaming chat, generative UI, chat history pe… | 62 | 1350 | active |
| ByteDance-Seed/SeedVR SeedVR/SeedVR2 are diffusion-transformer based models for generic real-world and AIGC video and image restoration, with SeedVR2 using adver… | 47 | 1347 | active |
| zanllp/infinite-image-browsing Infinite Image Browsing (IIB) is a full-featured image and video management application with fast thumbnail-based browsing, AI-generation m… | 92 | 1343 | active |
| galilai-group/lejepa LeJEPA is a Python framework for scalable, theoretically grounded self-supervised representation learning based on Joint-Embedding Predicti… | 45 | 1335 | active |
| moshstudio/TAICHI-flet TAICHI-flet is a Windows desktop entertainment application built with the Flet framework that lets users browse images, music, novels, comi… | 76 | 4733 | maintenance |
| LLaVA-VL/LLaVA-NeXT LLaVA-NeXT is a collection of open large multimodal models (LLaVA-NeXT, LLaVA-Video, LLaVA-OneVision, LLaVA-Critic-R1) that combine vision … | 64 | 4716 | maintenance |
| yukkcat/gemini-business2api A self-hosted gateway service that exposes Gemini Business through an OpenAI-compatible API, with multi-account load balancing and an admin… | 57 | 1324 | active |
| tensorflow/lucid Lucid is a collection of infrastructure and tools for research in neural network interpretability, built on TensorFlow 1.x. It provides fea… | 10 | 4702 | maintenance |
| meta-pytorch/segment-anything-fast A fast, batched offline inference-oriented fork of Meta's Segment Anything (SAM) image segmentation model. It applies optimizations like bf… | 44 | 1321 | active |
| yeyupiaoling/VoiceprintRecognition-Pytorch A PyTorch-based voiceprint recognition (speaker recognition) framework implementing models such as ECAPA-TDNN, ResNetSE, ERes2Net, and CAM+… | 57 | 1313 | active |
| sceneview/sceneview SceneView is a cross-platform 3D and augmented reality SDK built on Filament (Android/Web) and RealityKit (iOS), exposing declarative APIs … | 95 | 1303 | active |
| PantoMatrix/PantoMatrix PantoMatrix is an open-source research project that generates 3D face and body animation from speech audio, including the EMAGE model for h… | 33 | 1293 | active |
| AaronJackson/vrn Research code for the ICCV 2017 paper 'Large Pose 3D Face Reconstruction from a Single Image via Direct Volumetric CNN Regression'. It uses… | 32 | 4516 | maintenance |
| nianticlabs/monodepth2 Monodepth2 is the reference PyTorch implementation of the ICCV 2019 paper 'Digging into Self-Supervised Monocular Depth Prediction'. It tra… | 32 | 4500 | maintenance |
| NVlabs/stylegan2-ada-pytorch Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi… | 32 | 4487 | maintenance |
| luo3300612/Visualizer A lightweight Python library that extracts attention maps and other local variables from deep inside PyTorch models for visualization. It w… | 32 | 1270 | stable |
| selfxyz/self Self is an open-source monorepo for a privacy-preserving identity verification platform that lets users generate zero-knowledge proofs from… | 73 | 1259 | active |
| nv-tlabs/GET3D GET3D is NVIDIA's PyTorch implementation of a generative model that synthesizes high-quality 3D textured meshes (cars, chairs, animals, bui… | 32 | 4433 | maintenance |
| HelgeSverre/ollama-gui Ollama GUI is a web-based chat interface for interacting with local LLMs served through the Ollama API. It offers local chat history via In… | 70 | 1255 | active |
| ryokun6/ryos ryOS is a web-based desktop environment that recreates classic macOS and Windows interfaces in the browser, built with React and TypeScript… | 84 | 1244 | active |
| HJYao00/Mulberry Mulberry is a research implementation of an o1-like multimodal large language model (MLLM) that performs step-by-step reasoning and reflect… | 48 | 1243 | active |
| LTH14/fractalgen A PyTorch implementation of Fractal Generative Models (FractalGen), enabling pixel-by-pixel high-resolution image generation. It includes p… | 23 | 1243 | active |
| jcjohnson/fast-neural-style A Torch (Lua) implementation of feedforward neural style transfer from the ECCV 2016 paper 'Perceptual Losses for Real-Time Style Transfer … | 32 | 4360 | maintenance |
| xtreme1-io/xtreme1 Xtreme1 is an open-source, self-hosted data labeling and annotation platform for multimodal training data, supporting images, 3D LiDAR poin… | 62 | 1240 | active |
| facebookresearch/home-robot HomeRobot is an open-source robotics stack from Meta AI for mobile manipulation tasks on low-cost hardware like the Hello Robot Stretch. It… | 23 | 1236 | active |
| open-mmlab/playground OpenMMLab Playground is a central hub collecting and showcasing community projects that extend OpenMMLab libraries with Segment Anything Mo… | 30 | 1235 | active |
| ZED SDK The ZED SDK is a cross-platform spatial perception library for Stereolabs ZED stereo cameras, providing depth sensing, SLAM, 3D reconstruct… | 90 | 1228 | active |
| MotrixLab/SMPLer-X Official code for SMPLer-X, a family of foundation models for expressive human pose and shape estimation (EHPS) that unifies body, hand, an… | 59 | 1219 | stable |
| Taewan-P/gpt_mobile GPT Mobile is an Android chat application that lets users talk to multiple LLM providers (OpenAI, Anthropic, Google Gemini, Groq, Ollama) s… | 97 | 1218 | active |
| MyoHub/myosuite MyoSuite is a collection of musculoskeletal environments and tasks simulated with the MuJoCo physics engine and wrapped in the OpenAI gym A… | 92 | 1215 | active |
| OpenTSLM/OpenTSLM OpenTSLM is a family of Time-Series Language Models that integrate time series as a native modality into pretrained LLMs (Llama, Gemma), en… | 60 | 1213 | active |
| willisma/SiT Official PyTorch implementation of Scalable Interpolant Transformers (SiT), a family of generative models built on Diffusion Transformers t… | 52 | 1206 | active |
| toshas/torch-fidelity A PyTorch library providing accurate and efficient implementations of generative model evaluation metrics such as FID, Inception Score, KID… | 70 | 1200 | active |
| EvolvingLMMs-Lab/LLaVA-OneVision-2 A fully open framework for training multimodal large language models, releasing models, datasets, and training recipes for the LLaVA-OneVis… | 72 | 1197 | active |
| huggingface/llm.nvim llm.nvim is a Neovim plugin that brings LLM-powered features like ghost-text code completion to the editor, using llm-ls as its backend. It… | 67 | 1186 | active |
| AvenCores/open-antigravity-patcher An open-source Python patcher that removes regional restrictions from Google's Antigravity IDE, CLI, and VS Code extension, allowing use in… | 82 | 1185 | active |
| orpatashnik/StyleCLIP Official implementation of StyleCLIP, a method for text-driven manipulation of StyleGAN-generated imagery using CLIP. It provides three app… | 32 | 4120 | maintenance |
| mlfoundations/open_flamingo OpenFlamingo is an open-source PyTorch implementation of DeepMind's Flamingo, a large multimodal vision-language model that interleaves ima… | 23 | 4116 | maintenance |
| austin-weeks/miasma Miasma is a lightweight Rust web server that traps AI web scrapers in an endless pit of poisoned training data and self-referential links. … | 81 | 1181 | active |
| Abbey ScrapeServ is a self-hosted API service that accepts a URL and returns the website's data along with browser screenshots, using Playwright … | 24 | 1181 | active |
| warmshao/FasterLivePortrait A real-time portrait animation application based on LivePortrait that animates still photos or videos using a driving video, image, audio, … | 38 | 1178 | active |
| Tencent-Hunyuan/MixGRPO MixGRPO is a research framework from Tencent Hunyuan implementing a mixed ODE-SDE GRPO algorithm for efficient reinforcement learning fine-… | 58 | 1177 | active |
| csguoh/MambaIR MambaIR and MambaIRv2 are PyTorch-based image restoration models built on Mamba state-space models, published at ECCV 2024 and CVPR 2025. T… | 54 | 1174 | active |
| penxio/penx PenX is an open-source, local-first, privacy-first personal data hub for storing and managing structured data with AI-driven features. It i… | 50 | 1174 | active |
| chongzhou96/EdgeSAM EdgeSAM is the official PyTorch implementation of a distilled, accelerated variant of the Segment Anything Model (SAM) designed for on-devi… | 36 | 1174 | active |
| ztx888/HaloWebUI HaloWebUI is a self-hosted AI chat platform forked from Open WebUI, with a localized Chinese interface plus added model billing and usage s… | 54 | 1173 | active |
| bytedance/1d-tokenizer A research repository from ByteDance containing code and pretrained model weights for 1D visual tokenizers (TiTok, TA-TiTok, FlowTok) and i… | 29 | 1172 | active |
| magicleap/SuperGluePretrainedNetwork SuperGlue is a PyTorch implementation of a graph neural network with an optimal matching layer that matches sparse image features between t… | 32 | 4078 | maintenance |