A full-stack AI Red Teaming platform securing AI ecosystems via OpenClaw Security Scan, Agent Scan, Skills Scan, MCP scan, AI Infra scan and LLM jailbreak evaluation.
-
Updated
Aug 5, 2026 - Python
A full-stack AI Red Teaming platform securing AI ecosystems via OpenClaw Security Scan, Agent Scan, Skills Scan, MCP scan, AI Infra scan and LLM jailbreak evaluation.
The fastest Trust Layer for AI Agents
Leaderboard Comparing LLM Agent Security on System Prompt Leakage and Attack Probes
Extensible Go microkernel for LLM guardrails, quota enforcement and policy plugins.
Open-source prompt injection detector — 5 layers, 91.7% F1, ~27ms, offline, Apache 2.0
Universal Prompt Security Standard (UPSS): A framework for externalizing, securing, and managing LLM prompts and genAI systems, inspired by and extending OWASP OPSS concepts for any organization or project.
LLM Penetration Testing Framework - Discover vulnerabilities in AI applications before attackers do. 100attacks + AI-powered adaptive mode.
🚀 Unofficial Node.js SDK for Prompt Security's Protection API.
Mithra Scanner is an interactive API testing tool for prompt injection, refusal detection, and LLM security benchmarking. It supports YAML-based rule definitions, custom refusal lists, REST API integration, and provides detailed CLI output for security testing of language model endpoints.
CloakPrompt is a CLI tool that redacts secrets (passwords, API keys, credentials, etc.) before sending data to AI models.
The enterprise AI governance and safety control plane for AI agent prompts
Lightweight AI security framework for prompt validation, output scanning, risk scoring, sensitive data detection, and Zero Trust policy enforcement.
Local-first sanitizer for scripts, logs, prompts, and support text — clean secrets, hostnames, and org-specific terms on your device before sharing. PowerShell-aware Portfolio-code mode keeps sanitized code readable. Windows, Linux, and a live browser demo.
R package for LLM safety guardrails across prompts, outputs, RAG context, PII, secrets, and local Ollama/NLP workflows.
Industrial LLM agents, prompt safety, and orchestration
Static analysis CLI that scans codebases for LLM prompt-injection, data-exfiltration, jailbreak, and unsafe agent/tool vulnerabilities. Runs fully offline, integrates with CI/CD, and outputs console, JSON, and SARIF reports.
🛡️ Secure your LLM applications with PromptShields, a framework designed for real-time protection against prompt injection and data leaks.
Enterprise LLM prompt security & red-teaming framework — automated jailbreak detection, prompt injection testing, policy enforcement, and comprehensive audit trails for production AI systems.
Single-context metacognitive security framework for LLM prompt injection defense
Add a description, image, and links to the prompt-security topic page so that developers can more easily learn about it.
To associate your repository with the prompt-security topic, visit your repo's landing page and select "manage topics."