google-deepmind/pysc2
quality grade C, 51 out of 100StarCraft II Learning Environment
- stars
- 8.3k
- stars gained this week
- +3this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
RL algorithms, environments, simulators and decision-making systems.
Signals: reinforcement-learning, deep-reinforcement-learning, rl, gymnasium, openai-gym, multi-agent-reinforcement-learning, imitation-learning
663 results
StarCraft II Learning Environment
强化学习中文教程(蘑菇书🍄),在线阅读地址:https://datawhalechina.github.io/easy-rl/
Python Implementation of Reinforcement Learning: An Introduction
Machine learning, in numpy
Reinforcement Learning via Self-Distillation (SDPO)
[NeurIPS 2025] TTRL: Test-Time Reinforcement Learning
Understanding R1-Zero-Like Training: A Critical Perspective
Latest Advances on System-2 Reasoning
[ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
Scalable RL solution for advanced reasoning of language models
Implementation of all RL algorithms in a simpler way
MuZero
An elegant PyTorch deep reinforcement learning library.
Implementation of the paper "Towards Optimally Decentralized Multi-Robot Collision Avoidance via Deep Reinforcement Learning"
Dynamics and Domain Randomized Gait Modulation with Bezier Curves for Sim-to-Real Legged Locomotion.
A curated list of reinforcement learning with human feedback resources (continually updated)
Deep RL for MPC control of Quadruped Robot Locomotion
Mastering Atari with Discrete World Models
Sim-to-real RL training and deployment tools for the Unitree Go1 robot.
The reinforcement learning training code for AgiBot X1.
Imitation learning algorithms with Co-training for Mobile ALOHA: ACT, Diffusion Policy, VINN
Deep Reinforcement Learning for mobile robot navigation in ROS Gazebo simulator. Using Twin Delayed Deep Deterministic Policy Gradient (TD3) neural network, a robot learns to navigate to a random goal point in a simulated environment while avoiding obstacles.
Doosan robotic arm, simulation, control, visualization in Gazebo and ROS2 for Reinforcement Learning.
Deep Reinforcement Learning for Robotic Grasping from Octrees
24,535 repositories in the index in total.