Projects with this topic
-
Local-first, vision-LLM CAPTCHA solver for AI agents — MCP plugin + NopeCHA-style HTTP job API over any Chromium
Updated -
Official Ultralytics YOLO Flutter plugin for real-time inference on Android and iOS across major vision tasks.
Updated -
Ultralytics YOLO27, YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking
Updated -
YOLOE data pipeline for grounding and detection labels, predictions, text refinement, cache generation, and visualization.
Updated -
Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.
Updated -
PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
Updated -
Docker container, inference pipeline, and submission tooling for deploying trained YOLOv3 object detection models in the xView satellite imagery challenge.
Updated -
PyTorch sandbox for testing convolutional networks, ResNets, and other architectures on MNIST digits.
Updated -
YOLOv3 training, preprocessing, validation, and inference for object detection in xView satellite imagery and the xView detection challenge.
Updated -
An open-source computer vision framework for wildlife image analysis, featuring state-of-the-art models for species classification and detection.
Updated -
Semantic Vectorizing and Masking Creator: A desktop application for AI-powered image segmentation, masking, and vectorization.
Updated -
Turns ordinary photos into thermographic, X-ray, Kirlian, and night-vision renders via real physics simulation — not a filter.
Updated -
Aligns portrait photographs onto a fixed 2048 face grid for mosaic work
Updated -
Turn one crowded sheet of drawn icons into a clean, individually named, ready-to-use asset library. Local-first desktop app and CLI, with optional AI naming by a vision model running on your own machine.
Updated -
Master thesis: free and open-source hands-free cursor control through gaze and facial gestures, for people with severe motor paralysis.
Updated -
A pipeline for curating and sanitizing large-scale image datasets.
Updated -
Rendering and point-cloud generation toolkit for the PartNet and ShapeNet 3D shape datasets (ICCV 2021).
UpdatedUpdated -
Follow me - Vision is a vision-based control system designed to automate the navigation of remote-controlled objects like drones or cars. Using computer vision, the system detects a target object and generates commands to autonomously follow the subject in real time. Ideal for applications where a vehicle needs to track and follow a designated guide without manual intervention. For example, it could be used to help an elderly person by having a shopping cart follow them around a store.
This project is part of a larger, future initiative called Follow Me !
Updated -
Real-time YOLOv7/YOLOv7‑Tiny object detection on low‑powered Jetson Orin Nano using TensorRT and a multi‑threaded ROS 1 pipeline for edge robotics, drone monitoring.
Updated -
3D Hand Pose Estimation using Multi-View CNNs
Updated