X-GenGroup/Flow-Factory
quality grade B, 65 out of 100A unified framework for easy reinforcement learning in Flow-Matching models
- stars
- 649
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
RL algorithms, environments, simulators and decision-making systems.
Signals: reinforcement-learning, deep-reinforcement-learning, rl, gymnasium, openai-gym, multi-agent-reinforcement-learning, imitation-learning
653 results
A unified framework for easy reinforcement learning in Flow-Matching models
An Open Source package that allows video game creators, AI researchers and hobbyists the opportunity to learn complex behaviors for their Non Player Characters or agents
A Survey of Reinforcement Learning for Large Reasoning Models
A language agent gym with challenging scientific tasks
Projects from basic algorithms to MARL. Implements MADDPG,MATD3,MA/HAPPO in Predator-Prey pursuit games with PettingZoo MPE environments.
Turns your phone or VR headset into a robot arm teleoperation device by leveraging WebXR
We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference modeling and alignment problem.
Official code for "SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization"
A scalable asynchronous reinforcement learning implementation with in-flight weight updates.
A curated list of resources dedicated to reinforcement learning applied to cyber security.
A Type-1 Diabetes simulator implemented in Python for Reinforcement Learning purpose
This repo contains the Hugging Face Deep Reinforcement Learning Course.
Modular Reinforcement Learning (RL) library (implemented in PyTorch, JAX, and NVIDIA Warp) with support for Gymnasium/Gym, NVIDIA Isaac Lab, MuJoCo Playground and other environments
Modular Deep Reinforcement Learning framework in PyTorch. Companion library of the book "Foundations of Deep Reinforcement Learning".
Reinforcement Learning via Self-Distillation (SDPO)
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
HFTFramework utilized for research on " A reinforcement learning approach to improve the performance of the Avellaneda-Stoikov market-making algorithm "
[ECCV 2026] Official code of “MindDrive: A Vision-Language-Action Model for Autonomous Driving via Online Reinforcement Learning”
State of Health (SoH) and Remaining Useful Life (RUL) prediction for Li-ion batteries based on Physics-Informed Neural Networks (PINN).
Neural network learns to play snake in a terminal, built in Rust with Ratatui
High-speed simulator of convolutional spiking neural networks with at most one spike per neuron.
A long term short term memory recurrent neural network to predict forex data time series
Stock price prediction with recurrent neural network. The data is from the Chinese stock.
Generative adversarial networks (GAN) applied to sequential data via recurrent neural networks (RNN).
24,523 repositories in the index in total.