Showing 36 of 711 projects
A ROS-based system for robot localization and mapping using ceiling or floor-mounted fiducial markers.
A curated list of academic papers, datasets, and code for correcting rolling shutter effects, radial distortions, and text distortions in images and videos.
A curated list of resources for makeup and hairstyle transfer research using computer vision and generative AI.
A Python devkit for loading, exploring, and manipulating the PandaSet, a large-scale autonomous driving dataset with LiDAR, camera, and annotations.
An open-source tool to detect and blur faces and license plates in images for privacy compliance, using TensorFlow object detection.
A large-scale image dataset for self-supervised pretraining without humans, designed to reduce privacy concerns.
A PyTorch framework for deep learning on point clouds, providing a modular and reproducible foundation for 3D vision tasks.
A curated collection of resources on adversarial examples in deep learning, covering attacks, defenses, and applications.
A collection of pretrained deep learning models (StyleGAN2, GPT2, VGG, ResNet) for the Jax/Flax ecosystem.
Converts KITTI autonomous driving dataset raw data to ROS bags and provides a C++ library for direct data access.
An iOS library that uses face detection to calculate device distance and angle relative to a user's face for interactive 3D effects.
FLAME dataset and deep learning models for fire detection in aerial imagery using UAVs, supporting classification and segmentation tasks.
ROS package for sensor processing, object detection, tracking, and evaluation using the KITTI Vision Benchmark dataset.
A long-term autonomous driving dataset from Europe with multi-sensor data (GPS-RTK, LiDAR, cameras, IMU) for localization and mapping research.
A curated list of open-source software tools for medical imaging research, including segmentation, visualization, and deep learning libraries.
A Python library for interacting with FLIR thermal imaging cameras, capturing raw images, and converting proprietary file formats.
A Python library that simplifies using, finetuning, and deploying state-of-the-art machine learning models for various AI tasks.
A real-time object-level reconstruction system for 6D pose estimation using volumetric fusion and multi-object reasoning.
A ROS2 intelligent visual grasp solution for industrial robots, integrating OpenVINO grasp detection with MoveIt motion planning.
An iOS library that applies artistic styles to images using Core ML and pre-trained neural style transfer models.
ROS 2 packages for visual servoing and tracking using the ViSP library.
A Python-based CAPTCHA breaking solution using Keras and OpenCV, developed for a data science competition.
A large-scale driving behavior dataset with LiDAR point clouds, dashboard videos, and sensor data for autonomous driving research.
A curated collection of LiDAR place recognition methods, datasets, and algorithms for robotics and autonomous systems.
A deep learning model that reads IRCTC captchas with 98% accuracy, demonstrating their vulnerability to automated booking.
A foundation model for cell segmentation that achieves state-of-the-art performance across diverse cellular targets and imaging modalities.
A benchmark dataset and meta self-learning method for multi-source domain adaptation in scene text recognition.
A TensorFlow CNN implementation for Chinese character recognition, achieving 92.5% top-1 accuracy with batch normalization.
A vision transformer architecture that aggregates nested local transformers on image blocks for better accuracy, data efficiency, and convergence.
A single-header, zero-allocation C library for applying fast, chainable image filters compatible with SVG and CSS semantics.
A C++14 header-only library providing generic image representations and algorithms with performance close to hand-written code.
A scalable cell tracking method for 2D, 3D, and multichannel timelapse recordings, robust under segmentation uncertainty.
A CUDA-based implementation of KinectFusion for real-time dense surface reconstruction and tracking using a Kinect camera.
A tool for calibrating event cameras by converting event data to images and using standard image-based calibration toolboxes.
A simulation-based deep learning approach to enhance the resolution of 3D lidar point clouds for ground vehicles.
A PyTorch-based segmentation toolbox for electron microscopy connectomics, enabling neural structure analysis in 3D volumes.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.