Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A live stream development of RL tunning for LLM agents
| Date | Stars |
|---|---|
| 2026-07-31 | 4143 |
| 2026-08-02 | 4141 |
| 2026-08-03 | 4141 |
| 2026-08-06 | 4141 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# OpenManus-RL
🤗 <a href="https://huggingface.co/datasets/CharlieDreemur/OpenManus-RL" target="_blank">Dataset (OpenManus-RL)</a>
OpenManus-RL is an open-source initiative collaboratively led by __Ulab-UIUC__ and __MetaGPT__ .
This project is an extended version of the original [@OpenManus](https://github.com/FoundationAgents/OpenManus) initiative. Inspired by successful RL tunning for reasoning LLM such as Deepseek-R1, QwQ-32B, we will explore new paradigms for RL-based LLM agent tuning, particularly building upon foundations.
We are committed to regularly updating our exploration directions and results in a dynamic, live-streaming fashion. All progress, including rigorous testing on agent benchmarks such as GAIA, AgentBench, WebShop, and OSWorld, and tuned models, will be openly shared and continuously updated.
We warmly welcome contributions from the broader community—join us in pushing the boundaries of agent reasoning and tool integration!
Code and dataset are now available! The `verl` submodule has been integrated for enhanced RL training capabilities.
<div style="display: flex; justify-content: center;">
<div style="width: 100; transform: scale(1.0);">
<img src="assets/manus.jpg" style="width: 100%;" alt="marble">
</div>
</div>
## 📖 Table of Contents
- [OpenManus-RL](#openmanus-rl)
- [🔔 News](#-news)
- [Current Team Members](#current-team-members)
- [How to Contribute](#how-to-contribute)
- [Roadmap](#roadmap)
- [Method](#method)
- [Reasoning Models Exploration](#reasoning-models-exploration)
- [Alternative Rollout Strategies](#alternative-rollout-strategies)
- [Environment and Benchmark](#environment-and-benchmark)
- [Post-Training Strategies](#post-training-strategies)
- [Training of Agent Reward Model](#training-of-agent-reward-model)
- [Test-time Scaling of Trajectories](#test-time-scaling-of-trajectories)
- [Action Space Awareness and Strategic Exploration](#action-space-awareness-and-strategic-exploration)
- [Integration with RL Tuning Frameworks](#integration-with-rl-tuning-frameworks)
- [Dataset](#dataset)
- [Dataset Overbiew](#dataset-overview)
- [Data Instances](#data-instances)
- [Running](#Running)
- [Related Work](#related-work)
- [Agent tuning](#agent-tuning)
- [Tool using](#tool-using)
- [Agent tuning instruction dataset](#agent-tuning-instruction-dataset)
- [RL tuning](#rl-tuning)
- [Benchmark](#benchmark)
- [Similar Code](#similar-code)
- [Acknowledgement](#acknowledgement)
- [Community Group](#community-group)
- [Citation](#citation)
- [Documentation](#documentation)
---
## 🔔 News
- **[2025-03-09]** 🍺 We collect and opensource our Agent SFT dataset at [Huggingface](https://huggingface.co/datasets/CharlieDreemur/OpenManus-RL), go try it!
- **[2025-03-08]** 🎉 We are collaborating with [@OpenManus](https://github.com/mannaandpoem/OpenManus) from Metagpt to work on this project together!
- **[2025-03-06]** 🥳 We(UIUC-Ulab) are announcing our live-streaming project, OpenManus-RL.
## Current Team Members
[@Kunlun Zhu](https://github.com/Kunlun-Zhu)(Ulab-UIUC), [@Muxin Tian](https://github.com/realtmxi), [@Zijia Liu](https://m-serious.github.io/)(Ulab-UIUC), [@Yingxuan Yang](https://github.com/zoe-yyx),[@Jiayi Zhang](https://github.com/didiforgithub)(MetaGPT), [@Xinbing Liang](https://github.com/mannaandpoem), [@Weijia Zhang](https://github.com/CharlieDreemur), [@Haofei Yu](https://github.com/lwaekfjlk)(Ulab-UIUC), [@Cheng Qian](https://qiancheng0.github.io/),[@Bowen Jin](https://github.com/PeterGriffinJin),
---
# How to Contribute
We wholeheartedly welcome suggestions, feedback, and contributions from the community! Feel free to:
We welcome contributions, including fine-tuning codebase, tuning dataset, environment setup, and computing resources.
Create issues for feature requests, bug reports, or ideas.
Submit pull requests to help improve OpenManus-RL.
Or simply reach out to us for direct collaboration.
Important Excerpt of 18,903 characters
Read on GitHubKunlun Zhu
172
114
18
Manna Liang
14
4
3
Haofei Yu · CKC@ZJU -> LTI@CMU -> CS@UIUC · Israel
2
Zijia Liu
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:f6cb39fde47479c1, llm:description: 'A live stream development of RL tunning for LLM agents' (repository description)
matched fp:f6cb39fde47479c1, llm:description: 'A live stream development of RL tunning for LLM agents' (repository description)
matched fp:f6cb39fde47479c1, llm:description: 'A live stream development of RL tunning for LLM agents' (repository description)