agiresearch/ASB
quality grade C, 56 out of 100Agent Security Bench (ASB)
- stars
- 274
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Guardrails, PII redaction, prompt-injection defense, interpretability and offensive/defensive AI security.
Signals: ai-safety, ai-security, guardrails, prompt-injection, jailbreak, llm-security, interpretability, explainable-ai
381 results
Agent Security Bench (ASB)
SuperClaw: Red-Team AI Agents Before They Red-Team You
Keep your secrets hidden from AI agents.
Configurable Multi-layered AI Agentic Safety Framework
Security guard for AI agents — blocks malicious skills, prevents data leaks, protects secrets. 24 detection rules, runtime action evaluation, trust registry.
SlowMist Agent Security Skill: A comprehensive security review framework for AI agents operating in adversarial environments. Core principle: Every external input is untrusted until verified.
Jail your AI agent
SAF-MCP is a comprehensive security framework for documenting and mitigating threats in the AI Agent ecosystem.
Nova-Proximity is a MCP and Agent Skills security scanner powered with NOVA
Project CodeGuard is an AI model-agnostic security framework and ruleset that embeds secure-by-default practices into AI coding workflows (generation and review). It ships core security rules, translators for popular coding agents, and validators to test rule compliance.
Container-free, deny-by-default sandbox for AI coding agents. Kernel-enforced filesystem, network, and syscall isolation for Linux and macOS
A native policy enforcement layer for AI coding agents. Built on OPA/Rego.
Offensive security toolkit for Claude Code covering red team, exploit dev, AD attacks, EDR bypass, mobile pentest
Mirrored snapshot of Claude Code's source (exposed 2026-03-31) preserved for educational purposes, defensive security research, and software supply-chain analysis.
Lasso security integrations for Claude Code, including prompt-injection defenses
Claude Code skill for OWASP security best practices (2025-2026). Includes Top 10:2025, ASVS 5.0, Agentic AI security, and 20+ language-specific security quirks.
No description
No description
Claude Code 逆向工程研究仓库
SeClaw provides a practical foundation for measuring, diagnosing, and comparing security failures in autonomous LLM agents.
Powerful search page powered by LLMs and SearXNG
The automated prompt injection framework for LLM-integrated applications.
Redteaming LLMs using other LLMs
Dropbox LLM Security research code and results
24,523 repositories in the index in total.