vita-epfl/monoloco
quality grade D, 36 out of 100A 3D vision library from 2D keypoints: monocular and stereo 3D detection for humans, social distancing, and body orientation.
- stars
- 460
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Detection, segmentation, tracking, OCR models, 3D reconstruction and classical vision.
Signals: computer-vision, object-detection, image-segmentation, yolo, opencv, image-classification, pose-estimation, object-tracking
1,479 results
A 3D vision library from 2D keypoints: monocular and stereo 3D detection for humans, social distancing, and body orientation.
Implementing Vi(sion)T(transformer)
The official implementation of paper "ColorFlow: Retrieval-Augmented Image Sequence Colorization". ColorFlow:基于检索增强的图像序列上色
Instant-angelo: Build high-fidelity Digital Twin within 20 Minutes!
The Graph-Cut RANSAC algorithm proposed in paper: Daniel Barath and Jiri Matas; Graph-Cut RANSAC, Conference on Computer Vision and Pattern Recognition, 2018. It is available at http://openaccess.thecvf.com/content_cvpr_2018/papers/Barath_Graph-Cut_RANSAC_CVPR_2018_paper.pdf
CVPR 2022 papers with code (论文及代码)
Convert from Basel Face Model (BFM) to the FLAME head model
License Plate Detection using YOLOv8
Python Examples for Remote Sensing
[ICCV21] Self-Calibrating Neural Radiance Fields
This repo contains the updated version of all the assignments/labs (done by me) of Deep Learning Specialization on Coursera by Andrew Ng. It includes building various deep learning models from scratch and implementing them for object detection, facial recognition, autonomous driving, neural machine translation, trigger word detection, etc.
Tensorflow framework for the FLAME 3D head model. The code demonstrates how to sample 3D heads from the model, fit the model to 2D or 3D keypoints, and how to generate textured head meshes from Images.
Amazing Semantic Segmentation on Tensorflow && Keras (include FCN, UNet, SegNet, PSPNet, PAN, RefineNet, DeepLabV3, DeepLabV3+, DenseASPP, BiSegNet)
Code for PCN: Point Completion Network in 3DV'18 (Oral)
The First Place Solution of Kaggle iMaterialist (Fashion) 2019 at FGVC6
A Computer Vision based Traffic Signal Violation Detection System from video footage using YOLOv3 & Tkinter. (GUI Included)
SC-Depth (V1, V2, and V3) for Unsupervised Monocular Depth Estimation Webpage:https://jiawangbian.github.io/sc_depth_pl/
Code for Photo-Sketching: Inferring Contour Drawings from Images :dog:
[CVPR 2025, IJCV 2026] "A Distractor-Aware Memory for Visual Object Tracking with SAM2", "Distractor-Aware Memory-Based Visual Object Tracking"
[CVPR 2024✨Highlight] Official repository for HOLD, the first method that jointly reconstructs articulated hands and objects from monocular videos without assuming a pre-scanned object template and 3D hand-object training data.
Implementation of the KinectFusion approach in modern C++14 and CUDA
Python scripts performing object detection using the YOLOv8 model in ONNX.
[CVPR 2023] Official repository for downloading, processing, visualizing, and training models on the ARCTIC dataset.
Fashion Detection in the Wild (Deep Clothes Detector)
24,538 repositories in the index in total.