mll-lab-nu/VAGEN
quality grade B, 71 out of 100World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).
- stars
- 504
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
RL algorithms, environments, simulators and decision-making systems.
Signals: reinforcement-learning, deep-reinforcement-learning, rl, gymnasium, openai-gym, multi-agent-reinforcement-learning, imitation-learning
663 results
World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).
Uni-Agent is a framework for training long-horizon agents.
Awesome Deep Learning papers for industrial Search, Recommendation and Advertisement. They focus on Embedding, Matching, Pre-Ranking, Ranking, Post Ranking, Relevance, LLM and RL. Please cite our paper "Deep Learning to Rank in Industrial Search Engines, Recommender Systems, and Online Advertising - An Overview and New Perspectives" (TOIS 2026).
Drench yourself in Deep Learning, Reinforcement Learning, Machine Learning, Computer Vision, and NLP by learning from these exciting lectures!!
Monte Carlo tree search in JAX
Implementation of papers in 100 lines of code.
An engine for high performance multi-agent environments with very large numbers of agents, along with a set of reference environments
Uplift modeling and causal inference with machine learning algorithms
Simple and easily configurable grid world environments for reinforcement learning
UI-Venus is a general-purpose foundation GUI agent for mobile apps, web platforms, and desktop operating systems using only screenshots as input.
这是我的强化学习笔记
bsuite is a collection of carefully-designed experiments that investigate core capabilities of a reinforcement learning (RL) agent
TextWorld is a sandbox learning environment for the training and evaluation of reinforcement learning (RL) agents on text-based games.
A Next-Generation Training Engine Built for Ultra-Large MoE Models
A platform for Reasoning systems (Reinforcement Learning, Contextual Bandits, etc.)
This repo contains the Hugging Face Deep Reinforcement Learning Course.
High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)
A Survey of Reinforcement Learning for Large Reasoning Models
Modular Reinforcement Learning (RL) library (implemented in PyTorch, JAX, and NVIDIA Warp) with support for Gymnasium/Gym, NVIDIA Isaac Lab, MuJoCo Playground and other environments
Simulation of spiking neural networks (SNNs) using PyTorch.
PettingZoo and Gymnasium bindings for popular reinforcement learning environments outside of Farama
legged robot environments for reinforcement learning in multiple simulators (IsaacGym, Genesis, IsaacSim)
PyTorch Agent Net: reinforcement learning toolkit for pytorch
For deep RL and the future of AI.
24,535 repositories in the index in total.