Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
AgentLab: An open-source framework for developing, testing, and benchmarking web agents on diverse tasks, designed for scalability and reproducibility.
| Date | Stars |
|---|---|
| 2026-07-31 | 610 |
| 2026-08-05 | 615 |
| 2026-08-06 | 615 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<div align="center">
[](https://pypi.org/project/agentlab/)
[]([https://opensource.org/licenses/MIT](http://www.apache.org/licenses/LICENSE-2.0))
[](https://pypistats.org/packages/agentlab)
[](https://star-history.com/#ServiceNow/AgentLab)
[](https://github.com/ServiceNow/AgentLab/actions/workflows/code_format.yml)
[](https://github.com/ServiceNow/AgentLab/actions/workflows/unit_tests.yml)
[🛠️ Setup](#%EF%B8%8F-setup-agentlab) |
[🤖 Assistant](#-ui-assistant) |
[🚀 Launch Experiments](#-launch-experiments) |
[🔍 Analyse Results](#-analyse-results) |
<br>
[🏆 Leaderboard](#-leaderboard) |
[🤖 Build Your Agent](#-implement-a-new-agent) |
[↻ Reproducibility](#-reproducibility) |
[💪 BrowserGym](https://github.com/ServiceNow/BrowserGym)
<img src="https://github.com/user-attachments/assets/47a7c425-9763-46e5-be54-adac363be850" alt="agentlab-diagram" width="700"/>
[Demo solving tasks:](https://github.com/ServiceNow/BrowserGym/assets/26232819/e0bfc788-cc8e-44f1-b8c3-0d1114108b85)
</div>
> [!WARNING]
> AgentLab is meant to provide an open, easy-to-use and extensible framework to accelerate the field of web agent research.
> It is not meant to be a consumer product. Use with caution!
AgentLab is a framework for developing and evaluating agents on a variety of
[benchmarks](#-supported-benchmarks) supported by
[BrowserGym](https://github.com/ServiceNow/BrowserGym). It is presented in more details in our [BrowserGym ecosystem paper](https://arxiv.org/abs/2412.05467)
AgentLab Features:
* Easy large scale parallel [agent experiments](#-launch-experiments) using [ray](https://www.ray.io/)
* Building blocks for making agents over BrowserGym
* Unified LLM API for OpenRouter, OpenAI, Azure, or self-hosted using TGI.
* Preferred way for running benchmarks like WebArena
* Various [reproducibility features](#reproducibility-features)
* Unified [LeaderBoard](https://huggingface.co/spaces/ServiceNow/browsergym-leaderboard)
## 🎯 Supported Benchmarks
| Benchmark | Setup <br> Link | # Task <br> Template| Seed <br> Diversity | Max <br> Step | Multi-tab | Hosted Method | BrowserGym <br> Leaderboard |
|-----------|------------|---------|----------------|-----------|-----------|---------------|----------------------|
| [WebArena](https://webarena.dev/) | [setup](https://github.com/ServiceNow/BrowserGym/blob/main/browsergym/webarena/README.md) | 812 | None | 30 | yes | self hosted (docker) | soon |
| [WebArena-Verified](https://github.com/ServiceNow/webarena-verified) | [setup](https://github.com/ServiceNow/BrowserGym/blob/main/browsergym/webarena_verified/README.md) | 812 | None | 30 | yes | self hosted | soon |
| [WorkArena](https://github.com/ServiceNow/WorkArena) L1 | [setup](https://github.com/ServiceNow/WorkArena?tab=readme-ov-file#getting-started) | 33 | High | 30 | no | demo instance | soon |
| [WorkArena](https://github.com/ServiceNow/WorkArena) L2 | [setup](https://github.com/ServiceNow/WorkArena?tab=readme-ov-file#getting-started) | 341 | High | 50 | no | demo instance | soon |
| [WorkArena](https://github.com/ServiceNow/WorkArena) L3 | [setup](https://github.com/ServiceNow/WorkArena?tab=readme-ov-file#getting-started) | 341 | High | 50 | no | demo instance | soon |
| [WebLinx](https://mcgill-nlp.github.io/weblinx/) | - | 31586 | None | 1 | no | self hosted (dataset) | soon |
| [VisualWebArena](https://github.com/web-arena-x/visualwebarena) | [setup](https://giExcerpt of 17,228 characters
Read on GitHub283
268
185
Oleh Shliazhko
98
63
24
18
7
Xing Han Lu · @McGill-NLP · Canada
4
Gabriel Huang · @ServiceNow · Canada
2
2
Ikko Eltociear Ashimine · Japan
2
2
1
1
Xiangyi Li · @benchflow-ai · United States
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:248980d25024f495, topic:llm
matched fp:248980d25024f495, topic:benchmark
matched fp:248980d25024f495, topic:agents