opendilab/DI-sheep
quality grade D, 45 out of 100羊了个羊 + 深度强化学习(Deep Reinforcement Learning + 3 Tiles Game)
- stars
- 519
- stars gained this week
- —this week
- forks, open issues and contributors
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
RL algorithms, environments, simulators and decision-making systems.
Signals: reinforcement-learning, deep-reinforcement-learning, rl, gymnasium, openai-gym, multi-agent-reinforcement-learning, imitation-learning
663 results
羊了个羊 + 深度强化学习(Deep Reinforcement Learning + 3 Tiles Game)
The codes of paper "Long Text Generation via Adversarial Training with Leaked Information" on AAAI 2018. Text generation using GAN and Hierarchical Reinforcement Learning.
Implementation of: Nazari, Mohammadreza, et al. "Deep Reinforcement Learning for Solving the Vehicle Routing Problem." arXiv preprint arXiv:1802.04240 (2018).
[ICLR 2026] On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification.
DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.
[ICML 2026] Official resources of "Graph-R1: Towards Agentic GraphRAG Framework via End-to-end Reinforcement Learning".
Trading Gym is an open source project for the development of reinforcement learning algorithms in the context of trading.
Code for the paper "Offline Reinforcement Learning as One Big Sequence Modeling Problem"
Code for the paper "Training Diffusion Models with Reinforcement Learning"
This is a repository for reinforcement learning implementation for Unitree robots, based on Mujoco.
Reimplementation of DDPG(Continuous Control with Deep Reinforcement Learning) based on OpenAI Gym + Tensorflow
A parallel framework for population-based multi-agent reinforcement learning.
Playing Flappy Bird Using Deep Reinforcement Learning (Based on Deep Q Learning DQN using Tensorflow)
NeurIPS 2023: Safety-Gymnasium: A Unified Safe Reinforcement Learning Benchmark
Reinforcement Learning example in C, playing tic tac toe
Framework for Multi-Agent Deep Reinforcement Learning in Poker
Learning to trade under the reinforcement learning framework
Multi-Objective Reinforcement Learning algorithms implementations.
Cloud-native Financial Reinforcement Learning
Related papers for reinforcement learning, including classic papers and latest papers in top conferences
Repository for our papers: Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics and Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL algorithms, tasks, and models while being systematically grounded in its two core tenets: reproducibility and standardization.
This repository contains a collection of resources and papers on Diffusion Models for RL, accompanying the paper "Diffusion Models for Reinforcement Learning: A Survey"
This project uses reinforcement learning on stock market and agent tries to learn trading. The goal is to check if the agent can learn to read tape. The project is dedicated to hero in life great Jesse Livermore.
24,537 repositories in the index in total.