Ross ROSS = Recommend OSS · open-source software intelligence for agents

domain: artificial-intelligence

4539 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
facebookresearch/fast3r
Fast3R is the official PyTorch implementation of a CVPR 2025 model from Meta FAIR that reconstructs 3D scenes and estimates camera poses fr…
101591active
Korea Investment Open Trading API
Official sample code repository from Korea Investment & Securities for their Open Trading API (KIS Developers), providing Python examples f…
761590active
Drexubery/ViewCrafter
ViewCrafter is a research codebase that uses video diffusion models to synthesize high-fidelity novel views of scenes from a single or spar…
481590active
meta-pytorch/torchtune
Torchtune is a PyTorch-native library for authoring, post-training, and experimenting with large language models. It provides hackable trai…
705802maintenance
gcorso/DiffDock
DiffDock is a deep learning implementation of a diffusion generative model for molecular docking, predicting how small molecule ligands bin…
311570active
roryclear/clearcam
Clearcam is a self-hosted Python NVR that adds AI object detection, tracking, mobile notifications, and semantic search to any RTSP securit…
841568active
openilink/openilink-hub
OpeniLink Hub is a self-hosted, open-source message management platform and app marketplace for WeChat ClawBot (iLink protocol) bots, built…
721565active
feenkcom/gtoolkit
Glamorous Toolkit is the Moldable Development environment built on Pharo/Smalltalk, providing an integrated IDE with thousands of contextua…
961562active
Roncoo Education System
Roncoo Education (领课教育系统) is an open-source online education platform built with a Spring Cloud Alibaba microservices backend and Vue 3/Nux…
681546active
RLHFlow/RLHF-Reward-Modeling
A collection of training recipes for reward models used in RLHF, covering Bradley-Terry reward models, pairwise preference models, ArmoRM, …
331541active
ROMP
ROMP is a PyTorch-based library and pip-installable API (simple-romp) for real-time monocular multi-person 3D human mesh recovery, implemen…
231539stable
feremabraz/bloomberg-terminal
A Bloomberg Terminal clone built with Next.js 15, React 19, and TypeScript that provides real-time financial market data visualization in a…
501534active
BrokenSource/DepthFlow
DepthFlow is a free, open-source Python application and library that converts still images into 3D parallax effect videos using monocular d…
841528active
Tencent/TFace
TFace is a research platform from Tencent Youtu Lab for trusty face analysis, covering face recognition, face security (anti-spoofing), fac…
571523active
allenzren/open-pi-zero
An open-source re-implementation of the pi0 vision-language-action (VLA) model from Physical Intelligence, built on a pre-trained PaliGemma…
241523active
decoderesearch/SAELens
SAELens is a Python library for training sparse autoencoders (SAEs) on language model activations and analyzing them for mechanistic interp…
891521active
PrathamLearnsToCode/paper2code
An agent skill for coding agents like Claude Code that converts an arXiv paper URL into a working Python implementation. It generates citat…
481517active
zapdos-labs/unblink
Unblink is an AI-powered camera monitoring application that uses a vision language model (Qwen3-VL) to analyze camera frames, summarize act…
501507active
edmund-io/edmunds-claude-code
A Claude Code plugin providing 14 slash commands and 11 specialized AI agents for web development workflows. It scaffolds APIs, React compo…
381502active
pluralsh/plural
Plural is an enterprise Kubernetes management platform that provides fleet-scale GitOps deployments, infrastructure-as-code management, and…
951497active
CUT3R/CUT3R
CUT3R is the official PyTorch implementation of 'Continuous 3D Perception Model with Persistent State' (CVPR 2025 Oral), a stateful recurre…
381489active
fawney19/Aether
Aether is a self-hosted AI API gateway written in Rust that provides a unified entry point for Claude, OpenAI, Gemini, and their CLI client…
811462active
LocoMuJoCo
LocoMuJoCo is an imitation learning benchmark for whole-body locomotion control built on MuJoCo, featuring humanoid, quadruped, and musculo…
751454active
zsyOAOA/InvSR
InvSR is a Python research library implementing arbitrary-steps image super-resolution via diffusion inversion, leveraging pre-trained diff…
501452active
tianweiy/DMD2
DMD2 is the official PyTorch implementation of Improved Distribution Matching Distillation, a NeurIPS 2024 method that distills diffusion m…
281448active
writer/writer-framework
Writer Framework is an open-source Python framework for building data and AI applications with a drag-and-drop visual editor for the fronte…
761447active
aeon-toolkit/aeon
aeon is a scikit-learn compatible Python toolkit for machine learning on time series, covering classification, regression, clustering, fore…
871440active
neuralchen/SimSwap
SimSwap is a PyTorch-based face-swapping framework that performs arbitrary face swaps on images and videos using a single trained model. It…
235186maintenance
honeyandme/RAGQnASystem
A medical intelligent question-answering system combining knowledge-graph RAG with large language models, built on the DiseaseKG dataset wi…
621437active
YanjieZe/3D-Diffusion-Policy
3D Diffusion Policy (DP3) is a visual imitation learning algorithm that combines compact 3D point cloud representations with diffusion poli…
461437active
google/oss-fuzz-gen
A Google framework that uses large language models to automatically generate fuzz targets for real-world C/C++, Java, and Python projects, …
581434active
Francis-Rings/StableAnimator
StableAnimator is an end-to-end ID-preserving video diffusion framework that animates a reference human image according to a sequence of po…
401431active
zsyOAOA/ResShift
ResShift is an efficient diffusion model for image super-resolution that transfers between low- and high-resolution images by shifting resi…
611427active
acon96/home-llm
A Home Assistant custom integration plus fine-tuned small language models that let you control your smart home entirely with a local LLM, n…
861426active
apple/ml-aim
Apple's official repository for AIM (Autoregressive Image Models), providing code and pretrained checkpoints for AIMv1 and AIMv2 large visi…
411424active
jakobhoeg/nextjs-ollama-llm-ui
A fully-featured, ChatGPT-inspired web interface for chatting with local Ollama LLMs, built with Next.js, React, and Tailwind. It runs full…
301423active
WPeGPT
WPeGPT is an IDA Pro plugin that integrates LLM models (OpenAI, DeepSeek, or any OpenAI-compatible API) into binary analysis workflows. It …
741422active
SagiPolaczek/NeuralSVG
Official PyTorch implementation of NeuralSVG, an ICCV 2025 paper that generates layered, editable SVG vector graphics from text prompts. It…
461419active
dexmal/dexbotic
Dexbotic is an open-source PyTorch-based toolbox for developing Vision-Language-Action (VLA) models for embodied intelligence. It unifies p…
741417active
nv-tlabs/GEN3C
GEN3C is NVIDIA's research codebase for a generative video model that achieves precise camera control and temporal 3D consistency using a 3…
591414active
lucidrains/self-rewarding-lm-pytorch
A PyTorch library implementing the Self-Rewarding Language Model training framework from MetaAI, along with the SPIN training method. It pr…
161410active
open-gigaai/giga-world-policy
GigaWorld-Policy is a World Action Model (WAM) for robot policy learning that jointly models actions and future visual observations during …
581403active
yfeng95/PRNet
PRNet is a Python/TensorFlow implementation of the ECCV 2018 Position Map Regression Network for joint 3D face reconstruction and dense ali…
325014maintenance
Junyi42/monst3r
MonST3R is the official PyTorch implementation of an ICLR 2025 paper that estimates per-timestep geometry (pointmaps) from dynamic videos i…
361387active
yanx27/Pointnet_Pointnet2_pytorch
A pure PyTorch implementation of the PointNet and PointNet++ deep learning architectures for point cloud processing. It includes training a…
324945maintenance
OpenPPL/ppl.nn
PPLNN is a high-performance deep-learning inference engine written in C++ that runs ONNX models on x86 CPUs and NVIDIA GPUs, with a dedicat…
321367active
andyhuo520/aetherviz-master
AetherViz Master is an AI-powered interactive educational visualization tool that turns any teaching topic into an immersive 3D interactive…
501365active
79E/ChatGpt-Web
A commercially-viable ChatGPT web application built with React and TypeScript, featuring a full admin backend for managing users, tokens, p…
261363active
Sense-X/Co-DETR
Co-DETR is a PyTorch implementation of DETRs with Collaborative Hybrid Assignments Training, an ICCV 2023 object detection and instance seg…
321360stable
lxtGH/OMG-Seg
Official research codebase for OMG-Seg (CVPR 2024) and OMG-LLaVA (NeurIPS 2024), unified models for image-level, object-level, and pixel-le…
471354active
wyhuai/DDNM
DDNM is a Python research codebase implementing the Denoising Diffusion Null-Space Model for zero-shot image restoration, published as an I…
321351stable
emonney/QuickApp
QuickApp is an opinionated full-stack project template combining Angular 21 and ASP.NET Core 10 with pre-built authentication, authorizatio…
821350active
vercel-labs/gemini-chatbot
An open-source Next.js chatbot template powered by Google Gemini and the Vercel AI SDK, with streaming chat, generative UI, chat history pe…
621350active
ByteDance-Seed/SeedVR
SeedVR/SeedVR2 are diffusion-transformer based models for generic real-world and AIGC video and image restoration, with SeedVR2 using adver…
471347active
zanllp/infinite-image-browsing
Infinite Image Browsing (IIB) is a full-featured image and video management application with fast thumbnail-based browsing, AI-generation m…
921343active
galilai-group/lejepa
LeJEPA is a Python framework for scalable, theoretically grounded self-supervised representation learning based on Joint-Embedding Predicti…
451335active
moshstudio/TAICHI-flet
TAICHI-flet is a Windows desktop entertainment application built with the Flet framework that lets users browse images, music, novels, comi…
764733maintenance
LLaVA-VL/LLaVA-NeXT
LLaVA-NeXT is a collection of open large multimodal models (LLaVA-NeXT, LLaVA-Video, LLaVA-OneVision, LLaVA-Critic-R1) that combine vision …
644716maintenance
yukkcat/gemini-business2api
A self-hosted gateway service that exposes Gemini Business through an OpenAI-compatible API, with multi-account load balancing and an admin…
571324active
tensorflow/lucid
Lucid is a collection of infrastructure and tools for research in neural network interpretability, built on TensorFlow 1.x. It provides fea…
104702maintenance
meta-pytorch/segment-anything-fast
A fast, batched offline inference-oriented fork of Meta's Segment Anything (SAM) image segmentation model. It applies optimizations like bf…
441321active
yeyupiaoling/VoiceprintRecognition-Pytorch
A PyTorch-based voiceprint recognition (speaker recognition) framework implementing models such as ECAPA-TDNN, ResNetSE, ERes2Net, and CAM+…
571313active
sceneview/sceneview
SceneView is a cross-platform 3D and augmented reality SDK built on Filament (Android/Web) and RealityKit (iOS), exposing declarative APIs …
951303active
PantoMatrix/PantoMatrix
PantoMatrix is an open-source research project that generates 3D face and body animation from speech audio, including the EMAGE model for h…
331293active
AaronJackson/vrn
Research code for the ICCV 2017 paper 'Large Pose 3D Face Reconstruction from a Single Image via Direct Volumetric CNN Regression'. It uses…
324516maintenance
nianticlabs/monodepth2
Monodepth2 is the reference PyTorch implementation of the ICCV 2019 paper 'Digging into Self-Supervised Monocular Depth Prediction'. It tra…
324500maintenance
NVlabs/stylegan2-ada-pytorch
Official PyTorch implementation of StyleGAN2-ADA, a generative adversarial network with adaptive discriminator augmentation for training wi…
324487maintenance
luo3300612/Visualizer
A lightweight Python library that extracts attention maps and other local variables from deep inside PyTorch models for visualization. It w…
321270stable
selfxyz/self
Self is an open-source monorepo for a privacy-preserving identity verification platform that lets users generate zero-knowledge proofs from…
731259active
nv-tlabs/GET3D
GET3D is NVIDIA's PyTorch implementation of a generative model that synthesizes high-quality 3D textured meshes (cars, chairs, animals, bui…
324433maintenance
HelgeSverre/ollama-gui
Ollama GUI is a web-based chat interface for interacting with local LLMs served through the Ollama API. It offers local chat history via In…
701255active
ryokun6/ryos
ryOS is a web-based desktop environment that recreates classic macOS and Windows interfaces in the browser, built with React and TypeScript…
841244active
HJYao00/Mulberry
Mulberry is a research implementation of an o1-like multimodal large language model (MLLM) that performs step-by-step reasoning and reflect…
481243active
LTH14/fractalgen
A PyTorch implementation of Fractal Generative Models (FractalGen), enabling pixel-by-pixel high-resolution image generation. It includes p…
231243active
jcjohnson/fast-neural-style
A Torch (Lua) implementation of feedforward neural style transfer from the ECCV 2016 paper 'Perceptual Losses for Real-Time Style Transfer …
324360maintenance
xtreme1-io/xtreme1
Xtreme1 is an open-source, self-hosted data labeling and annotation platform for multimodal training data, supporting images, 3D LiDAR poin…
621240active
facebookresearch/home-robot
HomeRobot is an open-source robotics stack from Meta AI for mobile manipulation tasks on low-cost hardware like the Hello Robot Stretch. It…
231236active
open-mmlab/playground
OpenMMLab Playground is a central hub collecting and showcasing community projects that extend OpenMMLab libraries with Segment Anything Mo…
301235active
ZED SDK
The ZED SDK is a cross-platform spatial perception library for Stereolabs ZED stereo cameras, providing depth sensing, SLAM, 3D reconstruct…
901228active
MotrixLab/SMPLer-X
Official code for SMPLer-X, a family of foundation models for expressive human pose and shape estimation (EHPS) that unifies body, hand, an…
591219stable
Taewan-P/gpt_mobile
GPT Mobile is an Android chat application that lets users talk to multiple LLM providers (OpenAI, Anthropic, Google Gemini, Groq, Ollama) s…
971218active
MyoHub/myosuite
MyoSuite is a collection of musculoskeletal environments and tasks simulated with the MuJoCo physics engine and wrapped in the OpenAI gym A…
921215active
OpenTSLM/OpenTSLM
OpenTSLM is a family of Time-Series Language Models that integrate time series as a native modality into pretrained LLMs (Llama, Gemma), en…
601213active
willisma/SiT
Official PyTorch implementation of Scalable Interpolant Transformers (SiT), a family of generative models built on Diffusion Transformers t…
521206active
toshas/torch-fidelity
A PyTorch library providing accurate and efficient implementations of generative model evaluation metrics such as FID, Inception Score, KID…
701200active
EvolvingLMMs-Lab/LLaVA-OneVision-2
A fully open framework for training multimodal large language models, releasing models, datasets, and training recipes for the LLaVA-OneVis…
721197active
huggingface/llm.nvim
llm.nvim is a Neovim plugin that brings LLM-powered features like ghost-text code completion to the editor, using llm-ls as its backend. It…
671186active
AvenCores/open-antigravity-patcher
An open-source Python patcher that removes regional restrictions from Google's Antigravity IDE, CLI, and VS Code extension, allowing use in…
821185active
orpatashnik/StyleCLIP
Official implementation of StyleCLIP, a method for text-driven manipulation of StyleGAN-generated imagery using CLIP. It provides three app…
324120maintenance
mlfoundations/open_flamingo
OpenFlamingo is an open-source PyTorch implementation of DeepMind's Flamingo, a large multimodal vision-language model that interleaves ima…
234116maintenance
austin-weeks/miasma
Miasma is a lightweight Rust web server that traps AI web scrapers in an endless pit of poisoned training data and self-referential links. …
811181active
Abbey
ScrapeServ is a self-hosted API service that accepts a URL and returns the website's data along with browser screenshots, using Playwright …
241181active
warmshao/FasterLivePortrait
A real-time portrait animation application based on LivePortrait that animates still photos or videos using a driving video, image, audio, …
381178active
Tencent-Hunyuan/MixGRPO
MixGRPO is a research framework from Tencent Hunyuan implementing a mixed ODE-SDE GRPO algorithm for efficient reinforcement learning fine-…
581177active
csguoh/MambaIR
MambaIR and MambaIRv2 are PyTorch-based image restoration models built on Mamba state-space models, published at ECCV 2024 and CVPR 2025. T…
541174active
penxio/penx
PenX is an open-source, local-first, privacy-first personal data hub for storing and managing structured data with AI-driven features. It i…
501174active
chongzhou96/EdgeSAM
EdgeSAM is the official PyTorch implementation of a distilled, accelerated variant of the Segment Anything Model (SAM) designed for on-devi…
361174active
ztx888/HaloWebUI
HaloWebUI is a self-hosted AI chat platform forked from Open WebUI, with a localized Chinese interface plus added model billing and usage s…
541173active
bytedance/1d-tokenizer
A research repository from ByteDance containing code and pretrained model weights for 1D visual tokenizers (TiTok, TA-TiTok, FlowTok) and i…
291172active
magicleap/SuperGluePretrainedNetwork
SuperGlue is a PyTorch implementation of a graph neural network with an optimal matching layer that matches sparse image features between t…
324078maintenance

← prev page 43 / 46 next →