microsoft/BiomedParse
quality grade B, 71 out of 100BiomedParse: A Foundation Model for Joint Segmentation, Detection, and Recognition of Biomedical Objects Across Nine Modalities
- stars
- 707
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Detection, segmentation, tracking, OCR models, 3D reconstruction and classical vision.
Signals: computer-vision, object-detection, image-segmentation, yolo, opencv, image-classification, pose-estimation, object-tracking
1,479 results
BiomedParse: A Foundation Model for Joint Segmentation, Detection, and Recognition of Biomedical Objects Across Nine Modalities
fastdup is a powerful, free tool designed to rapidly generate valuable insights from image and video datasets. It helps enhance the quality of both images and labels, while significantly reducing data operation costs, all with unmatched scalability.
Techniques for deep learning with satellite & aerial imagery
[CVPR 2023 Highlight] Neural Kernel Surface Reconstruction
Local-first Video Knowledge Base. Index your video library with multi-modal analysis (YOLO, DeepFace, Whisper), search semantically via natural language, Docker-ready.
PyTorch implementation of YOLOv3, YOLOv3-SPP, and YOLOv3-tiny for real-time object detection with training, validation, inference, and multi-format export.
Claude Code skill for parametric 3D-printable model generation with CadQuery
Ultra-lightweight (600KB) Face Anti-Spoofing classifier. Optimized MiniFASNetV2-SE implementation validated on 70k+ samples with ~98% accuracy for edge devices.
OrbbecSDK python binding
A camera ISP (image signal processor) pipeline that contains modules with simple to complex algorithms implemented at the application level.
📸 Automatically detects and crops faces from batches of pictures.
Java interface to OpenCV, FFmpeg, and more
RNA-seq prediction with deep convolutional neural networks.
Earth observation processing framework for machine learning in Python
PyTorch3D is FAIR's library of reusable components for deep learning with 3D data
The collection of pre-trained, state-of-the-art AI models for ailia SDK
VisioFirm: Cross-Platform AI-assisted Annotation Tool for Computer Vision
Multiple Object Tracker, Based on Hungarian algorithm + Kalman filter.
Simba is a program used to repeat certain (complicated) tasks. Typically these tasks involve using the mouse and keyboard. Simba is programmable, which means you can design your own logic and steps that Simba will follow, based upon certain input such as colors on the screen.
🖼️A modern media gallery, with features like backup/sync, semantic search, media map, face recognition, memories and much more built using the latest Android technologies.
The NASA Vision Workbench is a general purpose image processing and computer vision library developed by the Autonomous Systems and Robotics (ASR) Area in the Intelligent Systems Division at the NASA Ames Research Center.
C-based/Cached/Core Computer Vision Library, A Modern Computer Vision Library
Installation support for Deep Learning Frameworks for the ArcGIS System
This repository is a paper digest of Transformer-related approaches in visual tracking tasks.
24,538 repositories in the index in total.