ybuild-ai/ai-game-art-pipeline-skill
quality grade B, 65 out of 100Agent skill from Y Build for turning AI images and videos into playable game art assets
- stars
- 313
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Detection, segmentation, tracking, OCR models, 3D reconstruction and classical vision.
Signals: computer-vision, object-detection, image-segmentation, yolo, opencv, image-classification, pose-estimation, object-tracking
1,478 results
Agent skill from Y Build for turning AI images and videos into playable game art assets
[Arxiv-2024] MotionLLM: Understanding Human Behaviors from Human Motions and Videos
[CVPR2026] TDMM-LM: Bridging Facial Understanding and Animation via Language Models
Contextual Object Detection with Multimodal Large Language Models
M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models
Learn to build and deploy local Visual Language Models for Edge AI
Read Like Humans: Autonomous, Bidirectional and Iterative Language Modeling for Scene Text Recognition
[ICCV'25 oral] Official Code for "LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models"
MonSter++: A Unified Geometric Foundation Model for Stereo and Multi-View Depth Estimation via the Unleashing of Monodepth Priors
Official repository for Dino U-Net: Exploiting High-Fidelity Dense Features from Foundation Models for Medical Image Segmentation. (DINOv3)
[CVPR 2024] Official implement of <Stronger, Fewer, & Superior: Harnessing Vision Foundation Models for Domain Generalized Semantic Segmentation>
[CVPR 2023] Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners
Official implementation of "Depth Any Panoramas: A Foundation Model for Panoramic Depth Estimation".
[CVPR 2024] Probing the 3D Awareness of Visual Foundation Models
[CVPR 2025] DEFOM-Stereo: Depth foundation model based stereo matching
[ECCV 2024] Official repository of Agent Attention
VIGA: Vision-as-Inverse-Graphics Agent
🖼️A modern media gallery, with features like backup/sync, semantic search, media map, face recognition, memories and much more built using the latest Android technologies.
A three.js agent skills for producing awesome graphics for scenes and games
Our method reconstructs 3D worlds from video diffusion models using non-rigid alignment to resolve inherent 3D inconsistencies in the generated sequences.
[CVPR'25 Oral] Official implementation for "DiffusionRenderer: Neural Inverse and Forward Rendering with Video Diffusion Models"
Official pytorch implementation for "LightenDiffusion: Unsupervised Low-Light Image Enhancement with Latent-Retinex Diffusion Models"
This is the official Pytorch implementation of the paper "Diffusion Models for Implicit Image Segmentation Ensembles".
[CVPR'24] Scaling Diffusion Models to Real-World 3D LiDAR Scene Completion
24,523 repositories in the index in total.