Sandbox any AI agent in seconds - zero setup, zero latency.
-
Updated
Aug 5, 2026 - Rust
Sandbox any AI agent in seconds - zero setup, zero latency.
Open-source AI agent firewall for MCP security and agent egress. Scans mediated HTTP, MCP, A2A, and WebSocket traffic for exfiltration, SSRF, and prompt injection, and emits mediator-signed action receipts: verifiable audit evidence from outside the agent.
Visibility into AI agent activity on endpoints, with on-device detection, optional pre-action blocking, and forensic reconstruction.
Open-source antivirus for AI agents: block risky tools, secret access, prompt injection, malicious packages, MCP servers, plugins, and skills at runtime.
Offline security scanner for AI-agent repos, skills, plugins, and MCP servers.
Security scanner MCP server for AI coding agents. Prompt injection firewall, package hallucination detection (4.3M+ packages), 1000+ vulnerability rules with AST & taint analysis, auto-fix.
Secure autonomous AI agent framework and platform. Build AI teams by describing what you want. Orchestrate agents that can do everything a human can do.
The deterministic merge gate for AI-generated agent capability changes — a local-first, static Tool-Use Readiness review for MCP, OpenAPI, and SDK tool surfaces. Open-source CLI + GitHub Action.
OpenAFW is a local agent firewall — keep your secrets off the model, the API relay, and the supply chain. Local credential masking, per-route model routing, and security detectors on the wire.
A curated timeline of real AI agent security incidents, breaches, and vulnerabilities (2024-2026). Every entry sourced and dated.
Hands-off supply-chain watchdog for dev machines: orchestrates multiple security scanners (Perplexity bumblebee + osv-scanner, govulncheck, NVIDIA SkillSpector) into one daily verdict — via Claude/Slack, desktop notification, or plain CLI.
Free OpenClaw security scanner. 3,000+ agents audited. 3-Layer Audit Protocol. OWASP ASI 10/10 coverage. AI agent integrity layer.
Six security outcomes an enterprise must achieve to run AI agents, and the controls that evidence them. A draft for public comment.
LLM guardrails & prompt injection detection for Python. Auto-instruments LangChain, CrewAI, OpenAI, LiteLLM + 8 more frameworks. PII masking, toxicity detection, policy CI/CD. One line, zero code changes.
Security-First Agentic IDE Coding Harness
25 production-tested defensive security skills for Claude Code - WordPress, VPS, Cloudflare, Next.js hardening, AI agent guardrails, MCP security, prompt injection defense, OWASP LLM Top 10, LLM coding failure modes (slopsquatting, hallucinated APIs, sycophancy), incident response, GDPR/DACH compliance. MIT, battle-tested.
Static scanner for MCP-connected AI agent pipelines — 271 rules across 12 categories, 12 compliance frameworks, OWASP Agentic 10/10 + MCP 10/10, GitHub Action, SARIF, public CVE-to-rule ledger.
AI got hands. This is the leash. Policy, audit, kill switch for any AI agent with access to your accounts.
Open security architecture for autonomous AI agents - extending Zero Trust principles
IntentFrame security plugin for Hermes Agent (Nous Research) — an external policy checkpoint that gates terminal, code, file, and cron tool calls before they run on your machine.
Add a description, image, and links to the ai-agent-security topic page so that developers can more easily learn about it.
To associate your repository with the ai-agent-security topic, visit your repo's landing page and select "manage topics."