System Prompt Leakage Detector MCP Connector for Claude
ADetects verbatim system prompt exfiltration using LCS algorithms.
This MCP server provides specialized security tools to identify if an AI agent has leaked its foundational instructions. By using exact Longest Common Substring (LCS) matching, it compares agent outputs against the original system prompt to calculate leakage percentages and identify specific character offsets of leaked text. It includes tools like detect_leakage to quantify risk, analyze_leakage_density to find clusters of leaked text, and get_risk_classification to map scores to severity levels.
Related Connectors
Acunetix 360 MCP
Automated web vulnerability scanning — manage scans, track issues, and audit security via AI.
Agent Error Propagation Tracker MCP
Trace error chains and calculate system impact in multi-agent environments.
Social Profile Visibility Check MCP
Audit social media profiles against privacy policies to find exposure mismatches.
Agent Loop Detector MCP
Identifies infinite conversation cycles and deadlocks in agentic workflows.