Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
A curated list of resources about AI agents for Computer Use, including research papers, projects, frameworks, and tools.
| Date | Stars |
|---|---|
| 2026-07-31 | 1719 |
| 2026-08-06 | 1722 |
Today
+3 stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
<div align="center">
<h1>
<div class="image-wrapper" style="display: inline-block;">
<img src="img/logo.png" alt="logo" height="100" style="display: block; margin: auto;">
</div>
[](https://x.com/i/communities/1874549355442802764)
ACU - Awesome Agents for Computer Use
</h1>
</div>
> An AI Agent for Computer Use is an autonomous program that can **reason** about tasks, **plan** sequences of actions, and **act** within the domain of a computer or mobile device in the form of clicks, keystrokes, other computer events, command-line operations and internal/external API calls. These agents combine perception, decision-making, and control capabilities to interact with digital interfaces and accomplish user-specified goals independently.
A curated list of resources about AI agents for Computer Use, including research papers, projects, frameworks, and tools.
## Table of Contents
- [ACU - Awesome Agents for Computer Use](#acu---awesome-agents-for-computer-use)
- [Table of Contents](#table-of-contents)
- [Articles](#articles)
- [Papers](#papers)
- [Surveys](#surveys)
- [Frameworks & Models](#frameworks--models)
- [UI Grounding](#ui-grounding)
- [Dataset](#dataset)
- [Benchmark](#benchmark)
- [Safety](#safety)
- [Projects](#projects)
- [Open Source](#open-source)
- [Frameworks & Models](#frameworks--models-1)
- [Environment & Sandbox](#environment--sandbox)
- [Automation](#automation)
- [Commercial](#commercial)
- [Frameworks & Models](#frameworks--models-2)
- [Contributing](#contributing)
## Articles
- [Anthropic | Introducing computer use, a new Claude 3.5 Sonnet, and Claude 3.5 Haiku](https://www.anthropic.com/news/3-5-models-and-computer-use)
- [Bill Gates | AI is about to completely change how you use computers](https://www.gatesnotes.com/AI-agents)
- [Ethan Mollick | When you give a Claude a mouse](https://www.oneusefulthing.org/p/when-you-give-a-claude-a-mouse)
- [OpenAI | Introducing Operator: A research preview of an agent that can use its own browser to perform tasks for you](https://openai.com/index/introducing-operator)
## Papers
<details open>
<summary><b>Surveys</b></summary>
### Surveys
- [AI Agents for Computer Use: A Review of Instruction-based Computer Control, GUI Automation, and Operator Assistants](https://arxiv.org/abs/2501.16150) (Jan. 2025)
- Comprehensive review establishing taxonomy of computer control agents (CCAs) from environment, interaction, and agent perspectives, analyzing 86 CCAs and 33 datasets
- [GUI Agents: A Survey](https://arxiv.org/abs/2412.13501) (Dec. 2024)
- General survey of GUI agents
- [Large Language Model-Brained GUI Agents: A Survey](https://arxiv.org/abs/2411.18279) (Nov. 2024)
- Focus on LLM-based approaches
- [Website](https://vyokky.github.io/LLM-Brained-GUI-Agents-Survey/)
- [GUI Agents with Foundation Models: A Comprehensive Survey](https://arxiv.org/abs/2411.04890) (Nov. 2024)
- Comprehensive overview of foundation model-based GUI agents
<br/>
</details>
<details open>
<summary><b>Frameworks & Models</b></summary>
### Frameworks & Models
- [Reinforcement Learning for Long-Horizon Interactive LLM Agents](https://arxiv.org/abs/2502.01600) (Feb. 2025)
- Novel RL approach (LOOP) for training IDAs directly in target environments
- 32B parameter agent outperforms OpenAI o1 by 9 percentage points on AppWorld
- [Large Action Models: From Inception to Implementation](https://arxiv.org/abs/2412.10047) (Dec. 2024)
- Comprehensive framework for developing LAMs that can perform real-world actions beyond language generation
- Details key stages including data collection, model training, environment integration, grounding and evaluation
- [Guiding VLM Agents with Process Rewards at Inference Time for GUI Navigation](https://openreview.net/forum?id=jR6YMxVG9i) (Dec. 2024)
- Novel reward-guided Excerpt of 26,362 characters
Read on GitHub20
2
2
Vardaan Pahuja
1
1
1
1
Nadeesha Cabral · Australia
1
ddupont · MIT
1
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:e4d972c59a6cb7ce, topic:computer-use, topic:gui-agent, desc:computer use
matched fp:e4d972c59a6cb7ce, topic:awesome, desc:curated list