Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
An awesome list of Agent Harness engineering resources, including GitHub projects, tools, benchmarks, and practical guides.
| Date | Stars |
|---|---|
| 2026-07-31 | 1535 |
| 2026-08-04 | 1554 |
| 2026-08-06 | 1554 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# Awesome Agent Harness A curated, implementation-first list of **agent harness engineering** resources, with GitHub projects as the primary focus. - Total entries: **338** - GitHub entries: **311 (92.0%)** - GitHub in project categories (excluding readings): **306/306 (100.0%)** - Categories: **9** - Last verified: **2026-06-21** - Language: [English](./README.md) | [中文](./README_zh.md) <a id="featured-harness-blogs"></a> ## Featured Harness Blogs - [Scaling Managed Agents: Decoupling the brain from the hands](https://www.anthropic.com/engineering/managed-agents): Anthropic's meta-harness architecture for decoupling session logs, harness loops, and sandboxes in long-horizon agents. - [What We Learned Building Cloud Agents](https://cognition.ai/blog/what-we-learned-building-cloud-agents): Cognition's field report on secure cloud-agent infrastructure, VM isolation, full-state snapshots, orchestration, governance, integrations, and enterprise adoption. - [Claude Code auto mode](https://www.anthropic.com/engineering/claude-code-auto-mode): Anthropic's write-up on classifier-backed approval delegation for safer high-autonomy coding-agent runs. - [Harness engineering (OpenAI)](https://openai.com/index/harness-engineering/): Field report on building reliable agent-first software via harness constraints and verification. - [The next evolution of the Agents SDK](https://openai.com/index/the-next-evolution-of-the-agents-sdk/): OpenAI's product and engineering post on model-native agent harnesses, native sandbox execution, manifests, memory, and filesystem and shell tools. - [Building Effective AI Agents](https://www.anthropic.com/engineering/building-effective-agents): Anthropic's practical guidance on when to use workflows vs. autonomous agents and how to structure them. - [Writing effective tools for AI agents](https://www.anthropic.com/engineering/writing-tools-for-agents): Best practices for tool interface design so agents call tools safely and reliably. - [Effective harnesses for long-running agents](https://www.anthropic.com/engineering/effective-harnesses-for-long-running-agents): Practical guide to maintaining state, resumability, and reliability over long agent runs. - [Harness design for long-running application development](https://www.anthropic.com/engineering/harness-design-long-running-apps): Follow-up article on improving long-running app generation through harness structure. - [Improving Deep Agents with harness engineering](https://blog.langchain.com/improving-deep-agents-with-harness-engineering/): Evidence that harness improvements alone can move benchmark performance. - [Evaluating Deep Agents: Our Learnings](https://blog.langchain.com/evaluating-deep-agents-our-learnings/): LangChain's practical lessons on evaluating stateful and long-horizon agents. - [Your Agent Needs a Harness, Not a Framework](https://www.inngest.com/blog/your-agent-needs-a-harness-not-a-framework): Argument for reliability-first infrastructure around agents instead of framework-only thinking. ## Contents - [Category Overview](#category-overview) - [Featured Harness Blogs](#featured-harness-blogs) - [Catalog](#catalog) - [Harness Architecture & Orchestration](#harness-architecture-orchestration) - [Context & Working-State Engineering](#context-working-state-engineering) - [Execution Substrates & Sandboxing](#execution-substrates-sandboxing) - [Protocols, Tool Interfaces & Agent Contracts](#protocols-tool-interfaces-agent-contracts) - [Evaluation Harnesses & Benchmarks](#evaluation-harnesses-benchmarks) - [Observability & Reliability Operations](#observability-reliability-operations) - [Guardrails, Security & Governance](#guardrails-security-governance) - [Reference Harness Implementations](#reference-harness-implementations) - [Essential Readings & Ecosystem Maps](#essential-readings-ecosystem-maps) - [Maintenance Notes](#maintenance-notes) - [Citation](#citation) ## Category Overview | Category | Entries | | --- | ---
Excerpt of 119,269 characters
Read on GitHub56
Would you bet a product on this? Bounded 0–100 and slow moving.
matched fp:b7bd2902eb730120, topic:context-engineering