Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
OpenClaw-RL: Train any agent simply by talking
| Date | Stars |
|---|---|
| 2026-07-24 | 5606 |
| 2026-07-25 | 5606 |
| 2026-07-28 | 5606 |
| 2026-07-30 | 5606 |
| 2026-07-31 | 5618 |
| 2026-08-06 | 5618 |
Today
— stars today
This week
+12 stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.21%/day
<div align="center">
<h1 align="center">
<img src="assets/spacer.png" alt="" width="23" height="40" align="absmiddle" />
OpenClaw-RL<!--
--><sup>
<img src="assets/clawistool.png" alt="Claw-RL logo" width="23" height="40" align="absmiddle" />
<sup>
</h1>
<p><b>Empowering OpenClaw with RL — Train a personalized agent simply by talking to it.</b></p>
<p><b>Scalable RL in real-world settings — Agentic RL for terminal, GUI, SWE, and tool-call settings.</b></p>
</div>
<p align="center">
<img src="https://img.shields.io/badge/⚡_Fully_Async-yellow?style=for-the-badge" alt="Fully Async" />
<img src="https://img.shields.io/badge/💰_Zero_API_or_Zero_GPU-blue?style=for-the-badge" alt="Zero API or Zero GPU" />
<img src="https://img.shields.io/badge/🤖_Personalized-success?style=for-the-badge" alt="Personalized" />
<img src="https://img.shields.io/badge/🛠️_Auto_Optimization-orange?style=for-the-badge" alt="Auto" />
<img src="https://img.shields.io/badge/💬_Language_Feedback-purple?style=for-the-badge" alt="Language Feedback" />
<img src="https://img.shields.io/badge/🧠_Hybrid_RL-red?style=for-the-badge" alt="Hybrid RL" />
<img src="https://img.shields.io/badge/🌍_Real_World_Agentic_RL-green?style=for-the-badge" alt="General Agentic RL" />
<br><br>
<a href="https://arxiv.org/abs/2603.10165"><img src="https://img.shields.io/badge/📄_Tech_Report-red?style=flat-square" alt="Tech Report" /></a>
<a href="https://yinjjiew.github.io/projects/openclawrl1"><img src="https://img.shields.io/badge/Blog-Page-blue?style=flat-square" alt="OpenClaw-RL Blog" /></a>
<a href="https://openclaw.ai"><img src="https://img.shields.io/badge/OpenClaw-Plugin-orange?style=flat-square" alt="OpenClaw Plugin" /></a>
<a href="https://github.com/THUDM/slime"><img src="https://img.shields.io/badge/Slime-Supported-purple?style=flat-square" alt="Slime Based" /></a>
<a href="https://thinkingmachines.ai/tinker/"><img src="https://img.shields.io/badge/Tinker-Supported-yellow?style=flat-square" alt="Tinker Supported" /></a>
<a href="LICENSE"><img src="https://img.shields.io/badge/License-Apache_2.0-green?style=flat-square" alt="License Apache 2.0" /></a>
</p>
<p align="center">
<video src="https://github.com/user-attachments/assets/a58aacad-3c1d-47aa-bbd1-cf8c5f36de6f" controls width="200"></video>
</p>
## 📰 News
- **[2026/4/15]** 🙌 We sincerely thank [Fireworks AI](https://fireworks.ai) for its generous support of this project, which has enabled more experiments and faster iteration.
- **[2026/4/11]** ✨ Qwen3.5-4B/9B/27B is supported now, both text and multi-modal!
- **[2026/4/4]** 👨👦👦 We support optimizing a single model based on feedback from a group of people.
- **[2026/3/25]** 🙌 We sincerely thank [Tinker](https://thinkingmachines.ai/tinker/) for its generous support of this project, which has enabled more experiments and faster iteration.
- **[2026/3/20]** 💻 You can use your own openclaw now, simply install [this extension](https://github.com/Gen-Verse/OpenClaw-RL/tree/main/extensions/rl-training-headers).
- **[2026/3/13]** ☁️ OpenClaw-RL now supports both local GPU and cloud ([Tinker](https://thinkingmachines.ai/tinker/)) deployment. Launch with [**one line of code**](#combinemethod) — Hybrid RL, OPD, and Binary RL all supported!
- **[2026/3/12]** ⚡ We support LoRA training now!
- **[2026/3/10]** 📃 We have released our [**Technical Report**](https://arxiv.org/abs/2603.10165)! 🏆 Ranked **#1** on [HuggingFace Daily Papers](https://huggingface.co/papers/2603.10165)!
- **[2026/3/10]** 🔥 Huge updates today! We released a [new combination method](./openclaw-combine), along with an [interesting evaluation](./openclaw-test) of these OpenClaw-RL methods. Track 2 is released too, featuring scalable RL implementations for general agent settings across [terminal](./terminal-rl), [GUI](./gui-rl), [SWE](./swe-rl), and [tool-call](./toolcall-rl) scenarios. We only focus on real-world settings!
- **[2026/3/3]** 🙌 Working with Excerpt of 21,061 characters
Read on GitHub215
7
Ling Yang
6
-.- · China
4
4
3
1
Junbo Niu · Peking University · China
1
HE ZHU · Peking University · China
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:a25d0b5f772dcd11, topic:rlhf, readme:lora