System Prompt Leakage Detector MCP Connector for Claude
ADetects verbatim leaks of system prompts within agent outputs using LCS algorithms.
The System Prompt Leakage Detector MCP server provides a specialized engine for identifying data exfiltration in AI agents. By utilizing the Longest Common Substring (LCS) algorithm via dynamic programming, it compares an agent's output against its original system instructions to find exact character-for-character reproductions. The tool calculates a leakage percentage, identifies precise character offsets of leaked segments, and computes a security risk score based on the presence of sensitive keywords like 'MANDATORY' or 'priority'. This is essential for developers building secure AI agents that must protect their underlying logic and instructions from being revealed to users.
Related Connectors
Acunetix 360 MCP
Automated web vulnerability scanning — manage scans, track issues, and audit security via AI.
Agent Error Propagation Tracker MCP
Trace error chains and calculate system impact in multi-agent environments.
Social Profile Visibility Check MCP
Audit social media profiles against privacy policies to find exposure mismatches.
Agent Loop Detector MCP
Identifies infinite conversation cycles and deadlocks in agentic workflows.