Showing 36 of 711 projects
A standalone open-source software solution for DIY video security with computer vision, local storage, and no monthly fees.
An open-source mobile app that describes photos using audio for blind and visually-impaired users.
A JAX/Flax implementation of the Fréchet Inception Distance (FID) metric for evaluating generative models.
A pre-trained deep learning model for image classification that identifies 1000 object classes using the Inception-ResNet-v2 architecture.
A C++ implementation of the PMBP algorithm combining PatchMatch and Belief Propagation for correspondence field estimation.
A ROS2 wrapper for the Movidius Neural Compute Stick (NCS) providing object classification and detection services for images and video streams.
Matlab code that converts RGBD images into CAD-like 3D models using surface prediction and POV-Ray rendering.
A TensorFlow-based GAN model that automatically adds color to black and white images.
A deep learning model that detects mitosis in breast cancer tumor cell images to assist in tumor proliferation scoring.
Computer vision software for FRC robots to detect targets and communicate with roboRIO via NetworkTables.
A deep learning model that classifies sports videos into 487 different sports activities using a 3D convolutional neural network.
Unofficial OpenCV binding for the D programming language, enabling computer vision development with D.
A Lua/Torch7 package for creating and manipulating edge-weighted graphs on images for segmentation and analysis.
A curated collection of academic papers, datasets, and metrics for topology-aware delineation in computer vision and medical imaging.
A curated list of the best production-ready Hugging Face models for NLP, vision, audio, and multimodal tasks.
A curated list of academic resources, datasets, and implementations for image harmonization research.
Crystal language bindings for FFmpeg to extract and process video frames from files and streams.
A Caffe-based deep learning model for face detection with pre-trained weights and training scripts.
A Caffe-based implementation of the LaMem model for scoring image memorability using convolutional neural networks.
Scripts to generate a dataset with static frames from the Arcade Learning Environment for Atari games.
An OCaml implementation of Mask R-CNN for object detection, segmentation, and classification using the Owl numerical library.
MicroPython binding for ESP-DL models enabling face detection, recognition, human detection, cat detection, and image classification on ESP32 devices.
A Directus extension that automatically tags images and extracts colors using the Imagga API.
Python implementation of scan unfolding for KITTI LiDAR data to create dense cylindrical projections without systematic discretization artifacts.
TensorFlow implementation of a generative adversarial network for video generation with scene dynamics.
A bilateral awareness network combining transformers and convolutions for semantic segmentation of very fine resolution urban scene images.
A ROS2 package that marks detected objects on a map during SLAM using object analytics.
A deep learning model for image classification that identifies 1000 object classes using the ResNet-50 architecture.
A lightweight desktop tool for annotating images with rectangles and contours for computer vision tasks.
A Facebook Messenger bot that analyzes images using Google Cloud Vision API to provide insights like face detection, text recognition, and content moderation.
A Capacitor plugin for Google ML Kit Vision, enabling on-device face detection and facial feature analysis in mobile apps.
Terraform templates for deploying IBM Maximo Visual Inspection Edge on IBM Cloud.
A React Native app using TensorFlow.js and MobileNet for real-time object detection on mobile devices.
A Blazor Server app that uses Azure Computer Vision to extract printed text from uploaded images.
A Torch7 package for creating and manipulating edge-weighted graphs on videos, enabling video segmentation and analysis.
A minimal JAX/Flax implementation of DETR with optimizations like Flash Attention and Sinkhorn solver.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.