rayleizhu/BiFormer
quality grade D, 43 out of 100[CVPR 2023] Official code release of our paper "BiFormer: Vision Transformer with Bi-Level Routing Attention"
- stars
- 584
- stars gained this week
- +1this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Detection, segmentation, tracking, OCR models, 3D reconstruction and classical vision.
Signals: computer-vision, object-detection, image-segmentation, yolo, opencv, image-classification, pose-estimation, object-tracking
1,479 results
[CVPR 2023] Official code release of our paper "BiFormer: Vision Transformer with Bi-Level Routing Attention"
A PyTorch Implementation of Neural IMage Assessment
AI Image Signal Processing and Computational Photography. Official library for NTIRE (CVPR) and AIM (ICCV/ECCV) Challenges. You will find Learned ISPs, RAW Restoration-Upsampling-Reconstruction, Image Enhancement, Bokeh rendering and more!
Photometric optimization code for creating the FLAME texture space and other applications
[CVPR 2024 - Oral] Matching 2D Images in 3D: Metric Relative Pose from Metric Correspondences
Lists of resources useful for my PhD in computer vision
FlyCV is a high-performance library for processing computer visual tasks.
NVIDIA Deep learning Dataset Synthesizer (NDDS)
Building Convolutional Neural Networks From Scratch using NumPy
MoCoGAN: Decomposing Motion and Content for Video Generation
A new algorithm for retrieving topological skeleton as a set of polylines from binary images
Official implementation of the paper Plan2Scene.
Self-Supervised Learning of 3D Human Pose using Multi-view Geometry (CVPR2019)
Image Acquisition Library for GenICam-based Machine Vision System
Computer vision based ML training data generation tool :rocket:
Easy & Modular Computer Vision Detectors, Trackers & SAM - Run YOLOv9,v8,v7,v6,v5,R,X in under 10 lines of code.
Efficiently Scaling Up Video Annotation with Crowdsourced Marketplaces. IJCV 2012
State of the art autonomous navigation scripts using Ai, Computer Vision, Lidar and GPS to control an arducopter based quad copter.
[ICCV 2025] Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models
Training repository for OpenPose
FBOW (Fast Bag of Words) is an extremmely optimized version of the DBow2/DBow3 libraries.
This is tensorflow implementation for paper "Deep Image Matting"
Build fully-functioning computer vision models with PyTorch
Arbitrary object tracking at 50-100 FPS with Fully Convolutional Siamese networks.
24,538 repositories in the index in total.