System Prompt Leakage Detector Connector for AI agents.
1 live capability
Prevent instruction theft and protect your proprietary agent logic.
Waiting for input…
Why people use System Prompt Leakage Detector
Stop instruction theft with System Prompt Leakage Detector
With this MCP, you automate that entire audit. The capability scans every response against your original prompt and flags exactly where the leak occurred, giving you instant visibility into security breaches.
What Vinkius changes
You get an automated way to catch instruction theft before it hits production.
Use it from Claude, ChatGPT, Cursor or another AI client you already have.
One account · 5,900+ Connectors
- Real-world use case 01
Testing new guardrails
A developer runs a batch of adversarial prompts through their agent and uses this MCP to see if any instructions slipped through.
- Real-world use case 02
Production monitoring
An automated pipeline checks live agent responses for any signs of instruction exfiltration.
- Real-world use case 03
Security auditing
A researcher uses the capability to verify that sensitive keywords in the system prompt aren't appearing in public outputs.
Complete set · 1capability
The complete System Prompt Leakage Detector capability set.
These are the exact actions your AI can choose when you ask it to work with System Prompt Leakage Detector.
01
1 capability in this set.
Part of 1 available through System Prompt Leakage Detector.
- 01 Capability
Detect prompt leakage
Scans an agent's response against your original instructions to find exact matches. It identifies the specific parts of your prompt that were leaked and calculates a risk score.
Set up in minutes
One URL. Then ask System Prompt Leakage Detector to work.
Claude and ChatGPT only need the Connector URL. Copy it once, add it in settings, and use System Prompt Leakage Detector from the conversation.
Choose your client
Live previewAdvanced clients IDE · CLI
Claude · Web + desktop
Connector URL · ready to paste
Streamable HTTPhttps://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp - Step 01
Open Connectors
In Claude Web or Claude Desktop, open Settings and choose Connectors.
- Step 02
Add the URL
Choose Add custom connector, name it System Prompt Leakage Detector, and paste the URL above.
- Step 03
Turn it on in chat
Select +, open Connectors, and enable System Prompt Leakage Detector for the conversation.
ChatGPT · Web + desktop
Connector URL · ready to paste
Streamable HTTPhttps://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp - Step 01
Open MCP settings
On desktop, open Settings and MCP servers. On web, open your workspace app or connector settings.
- Step 02
Add the URL
Choose Add server with Streamable HTTP, or create a custom MCP app, then paste the System Prompt Leakage Detector URL.
- Step 03
Save and start
Save the connection and enable System Prompt Leakage Detector in your conversation. Desktop may ask you to restart once.
Cursor · IDE configuration
Advanced setup
{
"mcpServers": {
"system-prompt-leakage-detector": {
"url": "https://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp"
}
}
} - Step 01
Open MCP Settings
Press Cmd+Shift+P (macOS) or Ctrl+Shift+P (Windows/Linux) → search "MCP Settings"
- Step 02
Add the server config
Paste the JSON configuration above into the mcp.json file that opens
- Step 03
Save the file
Cursor will automatically detect the new Connector
- Step 04
Start using System Prompt Leakage Detector
Open Agent mode in chat and ask: "Using System Prompt Leakage Detector, help me...". 1 tools available
VS Code Copilot · IDE configuration
Advanced setup
{
"mcpServers": {
"system-prompt-leakage-detector": {
"url": "https://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp"
}
}
} - Step 01
Create MCP config
Create a .vscode/mcp.json file in your project root
- Step 02
Add the server config
Paste the JSON configuration above
- Step 03
Enable Agent mode
Open GitHub Copilot Chat and switch to Agent mode using the dropdown
- Step 04
Start using System Prompt Leakage Detector
Ask Copilot: "Using System Prompt Leakage Detector, help me...". 1 tools available
Windsurf · IDE configuration
Advanced setup
{
"mcpServers": {
"system-prompt-leakage-detector": {
"url": "https://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp"
}
}
} - Step 01
Open MCP Settings
Go to Settings → MCP Configuration or press Cmd+Shift+P and search "MCP"
- Step 02
Add the server
Paste the JSON configuration above into mcp_config.json
- Step 03
Save and reload
Windsurf will detect the new server automatically
- Step 04
Start using System Prompt Leakage Detector
Open Cascade and ask: "Using System Prompt Leakage Detector, help me...". 1 tools available
Cline · IDE configuration
Advanced setup
{
"mcpServers": {
"system-prompt-leakage-detector": {
"url": "https://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp"
}
}
} - Step 01
Open Cline MCP Settings
Click the Connectors icon in the Cline sidebar panel
- Step 02
Add remote server
Click "Add Connector" and paste the configuration above
- Step 03
Enable the server
Toggle the server switch to ON
- Step 04
Start using System Prompt Leakage Detector
Ask Cline: "Using System Prompt Leakage Detector, help me...". 1 tools available
Claude Code · Terminal command
Advanced setup
claude mcp add system-prompt-leakage-detector --transport http "https://edge.vinkius.com/vk_preview_yMAsL4yWDRbK69R8LLTjB0GTLTzNpZnBPICtwOyQ/mcp" - Step 01
Install Claude Code
Run npm install -g @anthropic-ai/claude-code if not already installed
- Step 02
Add the Connector
Run the command above in your terminal
- Step 03
Verify the connection
Run claude mcp to list connected servers, or type /mcp inside a session
- Step 04
Start using System Prompt Leakage Detector
Ask Claude: "Using System Prompt Leakage Detector, show me...". 1 tools are ready
Where the request belongs
Work System Prompt Leakage Detector can move forward.
AI developers and security engineers who can't afford to have their proprietary logic leaked via prompt injection.
AI Engineer
Auditing agent responses during the testing phase of a new deployment.
Security Researcher
Testing the robustness of LLM guardrails against exfiltration attacks.
Product Manager
Ensuring that customer-facing agents remain within their intended operational boundaries.
Build the capability set
Add more capabilities.
Each Connector adds new actions and data without changing how you work.
Browse ConnectorsSystem Prompt Leakage Detector
Detects verbatim leaks of system prompts within agent outputs using LCS algorithms.
Prompt Injection Pattern Scanner
Scans user-supplied text for structural patterns associated with prompt-injection attempts.
Prompt Injection Detection Engine
Scans user inputs and retrieved documents for prompt injection attacks using static pattern matching.
Prompt Injection Pattern Detector
Identify and score malicious instruction overrides and system metadata extraction attempts in user text.
Prompt Injection Shield Prover
LLMs cannot distinguish system instructions from user input. This capability forces 5-layer injection defense analysis: intent isolation, privilege containment, indirect vector scanning, output sanitization, and scope enforcement. OWASP LLM Top 10 #1 compliance.
Prompt System Override Resistance Scorer
Quantify the structural integrity and resistance of LLM system prompts against manipulation.
Bring your own AI
Change the model, client or framework. Keep System Prompt Leakage Detector connected.
-
Claude -
ChatGPT -
Gemini -
Cursor -
VS Code -
Windsurf -
ZCode -
Cline -
Zed -
Continue -
Kiro -
Roo Code -
Zencoder -
Goose -
Void -
Augment Code -
Amp -
Qodo -
Tabnine -
Pieces -
Sourcegraph Cody -
JetBrains -
Warp -
Amazon Q -
Antigravity -
BoltAI -
Raycast -
Jan -
LM Studio -
AnythingLLM -
Open WebUI -
Msty -
Cherry Studio -
LibreChat -
TypingMind -
Chorus -
5ire -
n8n -
LangChain -
LlamaIndex -
CrewAI -
Vercel AI SDK
Before you connect
Questions about System Prompt Leakage Detector.
The practical details behind the request, access and result.
How can I use System Prompt Leakage Detector to protect my prompts?
It scans agent outputs for exact copies of your original instructions. This helps you catch when a user successfully uses prompt injection to reveal your internal logic.
Can System Prompt Leakage Detector find partial leaks?
The capability focuses on detecting verbatim, character-for-character matches using the LCS algorithm. It is designed to identify exact reproductions of your instructions.
Does System Prompt Leakage Detector help with prompt injection?
Yes, it detects when an injection attack succeeds in leaking your instructions. By identifying leaked segments, you can refine your guardrails.
How does System Prompt Leakage Detector calculate risk?
It looks for specific high-risk keywords like 'MANDATORY' within the leaked text. If these words appear in the output, the security risk score increases automatically.
Can I automate this with my AI client?
Yes, you can connect it to Claude or Cursor to audit responses automatically as part of your development or monitoring workflow.
How does the detection mechanism work?
The detect_prompt_leakage capability uses a deterministic Longest Common Substring (LCS) algorithm to find exact matches between the system prompt and the agent output, identifying precisely where instructions have been leaked.
What is a security risk score?
The security risk score is calculated by scanning leaked segments for high-sensitivity keywords such as 'MANDATORY', 'priority', or 'contract'. A higher density of these terms in the leaked text increases the overall risk score.
Can this capability detect partial leaks?
Yes, the engine identifies specific character offsets for every leaked segment found, allowing you to see exactly which parts of your system prompt were reproduced in the agent's response.
One connection away
Give your agent a direct line to System Prompt Leakage Detector.
Connect System Prompt Leakage Detector once. Keep it beside 5,900+ managed Connectors when the next task needs more.
Explore every Connector No credit card required · Free tier available