eloialonso/diamond
quality grade D, 49 out of 100DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.
- stars
- 2.1k
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
RL algorithms, environments, simulators and decision-making systems.
Signals: reinforcement-learning, deep-reinforcement-learning, rl, gymnasium, openai-gym, multi-agent-reinforcement-learning, imitation-learning
663 results
DIAMOND (DIffusion As a Model Of eNvironment Dreams) is a reinforcement learning agent trained in a diffusion world model. NeurIPS 2024 Spotlight.
Deep Reinforcement Learning: Zero to Hero!
Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch
A repo for data science related questions and answers
Solutions of Reinforcement Learning, An Introduction
PPO x Family DRL Tutorial Course(决策智能入门级公开课:8节课帮你盘清算法理论,理顺代码逻辑,玩转决策AI应用实践 )
A library of enterprise-grade AI agents designed to democratize artificial intelligence and provide free, open-source alternatives to overvalued Y Combinator startups.
TradeMaster is an open-source platform for quantitative trading empowered by reinforcement learning :fire: :zap: :rainbow:
Awesome free machine learning and AI courses with video lectures.
Implementations of basic RL algorithms with minimal lines of codes! (pytorch based)
Clean, Robust, and Unified PyTorch implementation of popular Deep Reinforcement Learning (DRL) algorithms (Q-learning, Duel DDQN, PER, C51, Noisy DQN, PPO, DDPG, TD3, SAC, ASL)
A high-performance distributed training framework for Reinforcement Learning
The next generation deep reinforcement learning tookit
Debugging, monitoring and visualization for Python Machine Learning and Data Science
Reinforcement Learning / AI Bots in Card (Poker) Games - Blackjack, Leduc, Texas, DouDizhu, Mahjong, UNO.
OpenDILab Decision AI Engine. The Most Comprehensive Reinforcement Learning Framework B.P.
Minimal and Clean Reinforcement Learning Examples
PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinforcement learning using Kronecker-factored approximation (ACKTR) and Generative Adversarial Imitation Learning (GAIL).
Massively Parallel Deep Reinforcement Learning. 🔥
[ICML 2021] DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning | 斗地主AI
Learn Deep Reinforcement Learning in 60 days! Lectures & Code in Python. Reinforcement Learning + Deep Learning
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
Repo for the Deep Reinforcement Learning Nanodegree program
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
24,535 repositories in the index in total.