Ross ROSS = Recommend OSS · open-source software intelligence for agents

function: computer-vision

1555 products, primary matches first, then adoption-weighted; health v2 shown.

ProductHealth v2StarsMaturity
leftthomas/SRGAN
A PyTorch implementation of SRGAN, the CVPR 2017 generative adversarial network for photo-realistic single-image super-resolution. It inclu…
321248maintenance
lfz/DSB2017
The winning solution of team 'grt123' for the 2017 Data Science Bowl (DSB2017), a deep learning pipeline for detecting lung cancer from CT …
321241maintenance
BR-IDL/PaddleViT
PaddleViT is a collection of state-of-the-art Vision Transformer and MLP model implementations for PaddlePaddle 2.1+, covering image classi…
231238maintenance
openseg-group/openseg.pytorch
Official PyTorch implementations of semantic segmentation models OCNet, OCRNet, and SegFix, achieving state-of-the-art results on benchmark…
231237maintenance
jfzhang95/pytorch-video-recognition
A PyTorch library implementing C3D, R3D, and R2Plus1D models for video action recognition, with training scripts for UCF101 and HMDB51 data…
321236maintenance
ml4a/ml4a-ofx
A collection of openFrameworks applications in C++ for real-time interactive machine learning, aimed at artists and creative coders. It inc…
231223maintenance
facebookresearch/House3D
House3D is a virtual 3D environment of over 45k fully annotated indoor scenes from the SUNCG dataset, built for training embodied AI agents…
101200maintenance
VITA-Group/DeblurGANv2
Official PyTorch implementation of DeblurGAN-v2, an ICCV 2019 relativistic conditional GAN for single-image motion deblurring with a Featur…
321192maintenance
aitorzip/DeepGTAV
DeepGTAV is a C++ plugin for Grand Theft Auto V that converts the game into a vision-based self-driving car research environment. It expose…
321186maintenance
cpvrlab/ImagePlay
ImagePlay is an open-source desktop application for rapid prototyping of image processing algorithms, combining over 70 individual image pr…
231173maintenance
Jinnrry/RobotHelper
RobotHelper is an Android automation script framework written in Java, providing common building blocks like screen capture, image-based po…
231136maintenance
yu-takagi/StableDiffusionReconstruction
Research codebase reproducing Takagi and Nishimoto's CVPR 2023 method for reconstructing images a person viewed from fMRI brain activity us…
301127maintenance
snap-research/EfficientFormer
A PyTorch implementation of EfficientFormer and EfficientFormerV2, efficient vision transformer model families designed to run at MobileNet…
321116maintenance
Res2Net/Res2Net-PretrainedModels
Official PyTorch implementation of Res2Net, a multi-scale CNN backbone architecture published in TPAMI, with ImageNet-pretrained model weig…
321114maintenance
andyzeng/visual-pushing-grasping
PyTorch reference implementation of Visual Pushing and Grasping (VPG), which trains robotic agents via self-supervised deep reinforcement l…
321109maintenance
qiucheng025/zao-
A Python deep learning tool that identifies and swaps faces in images and videos, with extract, train, and convert workflows plus an option…
321104maintenance
PaddlePaddle/Paddle.js
Paddle.js is a browser-based deep learning inference engine for Baidu PaddlePaddle models, running via WebGL, WebGPU, or WebAssembly backen…
231103maintenance
fyu/drn
A PyTorch library implementing Dilated Residual Networks (DRN), which combine dilated convolutions with residual networks for image classif…
321102maintenance
jacobgil/vit-explain
A PyTorch library implementing explainability methods for Vision Transformers, including Attention Rollout and Gradient Attention Rollout. …
321098maintenance
maelfabien/Multimodal-Emotion-Recognition
A real-time multimodal emotion recognition web app built with Flask that analyzes emotions from text, audio, and video inputs using deep le…
321089maintenance
ethanhe42/channel-pruning
Reference implementation of the ICCV 2017 channel pruning method for accelerating very deep convolutional neural networks, using LASSO regr…
231088maintenance
deepmedic/deepmedic
DeepMedic is an efficient multi-scale 3D convolutional neural network for segmenting 3D medical scans such as MRI and CT. It is a Python-ba…
231063maintenance
HRNet/HRNet-Image-Classification
Official PyTorch implementation and training code for HRNet (High-Resolution Network) image classification models on ImageNet. It provides …
231056maintenance
piergiaj/pytorch-i3d
A PyTorch port of DeepMind's I3D (Inflated 3D ConvNet) models pretrained on the Kinetics dataset for video action recognition. It includes …
321054maintenance
4uiiurz1/pytorch-nested-unet
A PyTorch implementation of the UNet++ (Nested U-Net) architecture for image segmentation, based on the paper 'UNet++: A Nested U-Net Archi…
321045maintenance
goberoi/faceit
A Python script that simplifies swapping faces in videos using the deepfakes/faceswap library, with training data sourced from YouTube vide…
321032maintenance
qubvel/ttach
TTAch is a Python library for image test time augmentation (TTA) with PyTorch. It wraps existing models to apply augmentations like flips, …
231030maintenance
Gumpest/YOLOv5-Multibackbone-Compression
A YOLOv5-based toolbox for swapping in lightweight or high-accuracy backbones (TPH-YOLOv5, GhostNet, ShuffleNetV2, MobileNetV3-Small, Effic…
321020maintenance
snap-research/NeROIC
Official PyTorch implementation of NeROIC, a neural method for capturing 3D object geometry and material from online image collections and …
321012maintenance
xiaoyufenfei/Efficient-Segmentation-Networks
A PyTorch reference implementation collection of lightweight, real-time semantic segmentation models such as ENet, ERFNet, LEDNet, Fast-SCN…
321008maintenance
dimensionalOS/dimos
DimOS is a Python SDK and agent-native operating system for generalist robotics, letting users command humanoids, quadrupeds, and drones in…
864431experimental
cbh123/narrator
A Python app that watches your webcam and generates David Attenborough-style narration of what it sees, using GPT vision models and ElevenL…
694426experimental
IliasHad/edit-mind
Edit Mind is a local-first video knowledge base that indexes video libraries with multi-modal AI analysis (Whisper transcription, YOLO obje…
761789experimental
jlsutherland/doc2text
doc2text is a Python library that extracts high-quality text from poorly scanned PDFs by correcting resolution, cropping, and skew before O…
321278experimental
ddupont808/GPT-4V-Act
GPT-4V-Act is a multimodal AI agent that combines GPT-4V(ision) with a Chromium browser, using Set-of-Mark prompting and a DOM auto-labeler…
271058experimental
fchollet/deep-learning-models
A deprecated collection of Keras code and pre-trained weights for popular deep learning image classification models such as VGG16, VGG19, R…
237348abandoned
karpathy/neuraltalk
NeuralTalk is a Python+numpy implementation of Multimodal Recurrent Neural Networks that generate natural-language descriptions of images. …
325503abandoned
liuzhuang13/DenseNet
Reference implementation of DenseNet (Densely Connected Convolutional Networks), the CVPR 2017 Best Paper Award-winning CNN architecture, w…
324869abandoned
Greenwolf/social_mapper
Social Mapper is a Python 3 OSINT tool that enumerates and correlates social media profiles across sites like LinkedIn, Facebook, Twitter, …
324073abandoned
udacity/self-driving-car-sim
A Unity-based self-driving car simulator built for Udacity's Self-Driving Car Nanodegree, used to train cars to navigate road courses with …
103985abandoned
pavelgonchar/colornet
Colornet is a Python/TensorFlow neural network implementation that colorizes grayscale images, based on VGG16 features and YUV color-channe…
323552abandoned
LeeJunHyun/Image_Segmentation
A PyTorch implementation of four U-Net variants for image segmentation: U-Net, R2U-Net, Attention U-Net, and Attention R2U-Net. It includes…
323101abandoned
ryanjay0/miles-deep
Miles Deep is a C++ deep learning application built on Caffe that classifies each second of a pornographic video into six sexual act catego…
232658abandoned
jakeret/tf_unet
A generic U-Net implementation built on TensorFlow 1.x for training image segmentation models on arbitrary imaging data. Originally develop…
231907abandoned
sikuli/sikuli
Sikuli is a visual GUI automation tool that uses screenshot matching to find and interact with on-screen elements, originally developed as …
321727abandoned
ramprs/grad-cam
Official Torch (Lua) implementation of Grad-CAM, the ICCV 2017 gradient-weighted class activation mapping technique for producing visual ex…
321668abandoned
Zehaos/MobileNet
A TensorFlow implementation of Google's MobileNets, efficient convolutional neural networks for mobile vision applications, including Image…
321659abandoned
bonlime/keras-deeplab-v3-plus
A Keras implementation of the DeepLab v3+ semantic image segmentation model with pretrained weights imported from the original TensorFlow c…
231375abandoned
MIC-DKFZ/medicaldetectiontoolkit
A PyTorch framework providing 2D and 3D implementations of object detectors like Mask R-CNN, Retina Net, and Retina U-Net, tailored for med…
321357abandoned
piiswrong/deep3d
Deep3D is a CNN-based research project that automatically converts 2D images and videos into 3D by estimating per-pixel depth maps and gene…
321299abandoned
shekkizh/FCN.tensorflow
A TensorFlow implementation of Fully Convolutional Networks (FCN) for semantic segmentation, based on the reference code from the original …
321248abandoned
Teaonly/android-eye
An Android app that turns an old phone into a surveillance security camera, streaming H.264 video and G.726 audio over a built-in web serve…
101085abandoned
alexgkendall/caffe-segnet
A modified version of the Caffe deep learning framework implementing SegNet, a deep convolutional encoder-decoder architecture for semantic…
321083abandoned
DeepLearningKit/DeepLearningKit
DeepLearningKit is an open-source deep learning framework for Apple's iOS, OS X and tvOS, written in Swift and using Metal for GPU-accelera…
321059abandoned
KevinGong2013/ChineseIDCardOCR
A deprecated Swift library for optical character recognition of Chinese second-generation ID cards on iOS, using Vision and CoreML. It has …
321025abandoned

← prev page 16 / 16