Use Prompt Cache Hit Calculator with your AI.
Connect your account once and let the AI you already use work with it, without building another integration. Analyze prompt prefix caching performance, efficiency, and cost savings.
Developed, maintained, and hosted by Vinkius.
MCP VERIFIED · PRODUCTION READY · VINKIUS GUARANTEED
Waiting for input…
Works with modern AI clients that support MCP, including ChatGPT, Claude, Cursor, and more.
Complete set · 3 capabilities
The complete Prompt Cache Hit Calculator capability set.
These are the exact actions your AI can choose when you ask it to work with Prompt Cache Hit Calculator.
01-03
3 capabilities in this set.
Part of 3 available through Prompt Cache Hit Calculator.
- 01
Analyze cache performance
Provides a high-level overview of how well the cache is performing regarding hits, efficiency, and cost savings
- 02
Evaluate cache optimization
Identifies the ideal cache capacity and the degree of prefix overlap to guide infrastructure scaling
- 03
Inspect cache dynamics
Investigates the frequency of cache turnover and the specific overlap between individual requests
Observed, not estimated
836ms average. Fast in production.
Prompt Cache Hit Calculator is checked daily against the live service.
- Fastest day
- 642ms
- Slowest day
- 970ms
- 14-day trend
- Slowing+27%
Connect your client
One URL. Every client.
Activate the Connector, copy your link, and paste it into the client you already use. 3 capabilities arrive ready to run.
Preview access · not provider authentication
The vk_preview_* token belongs to Vinkius preview infrastructure. It lets Claude discover and display the capabilities of Prompt Cache Hit Calculator, so you can see the experience inside your AI.
It does not authenticate your account with Prompt Cache Hit Calculator. Actions requiring credentials or live account data may not run until you activate the Connector and authorize the service.
Prompt Cache Hit Calculator Connector
You're all set. Choose your MCP client and follow the setup instructions.
https://edge.vinkius.com/vk_preview_RtPMDdAVBYJh5FuqI4pTMxgwoz4sOxqI1BctS0MI/mcpClaude Desktop
Follow the steps below to connect in seconds.
- 1In Claude Desktop, open Settings → Connectors.
- 2Click “Add custom connector” and paste the connector link above as the remote MCP server URL.
- 3Click Add and start a new chat — Prompt Cache Hit Calculator capabilities are ready to use.
{
"mcpServers": {
"prompt-cache-hit-calculator-mcp": {
"url": "https://edge.vinkius.com/vk_preview_RtPMDdAVBYJh5FuqI4pTMxgwoz4sOxqI1BctS0MI/mcp"
}
}
}
Claude
ChatGPT
Cursor
VS Code
Windsurf
Claude Code
JetBrains
Cline
Step-by-step instructions for each client are in the guide. How to connect
FAQ
Questions Prompt Cache Hit Calculator owners ask.
- 01
How do I calculate the monetary value of my cache hits?
You can use the analyze_cache_performance capability. By providing the costPerToken parameter, the capability calculates the total cacheHitValue based on the tokens saved during successful hits.
- 02
What determines the optimal cache size?
The evaluate_cache_optimization capability determines the optimal cache size by calculating the 95th percentile of prefix lengths from your request logs.
- 03
How can I see if my cache is evicting too many items?
Use the inspect_cache_dynamics capability. It provides the cacheEvictionRate, which is the number of evictions divided by the total cache capacity.
Explore
More in Analytics
Prompt Cache Hit Rate Calculator AI Connector
Analyze LLM prompt prefix caching efficiency and performance.
ViewAI Response Caching ROI Calculator AI Connector
Calculate the financial impact and payback period of AI response caching.
ViewSpeculative Decoding Calculator AI Connector
Optimize LLM inference speed and cost using deterministic speculative decoding metrics.
ViewAgent Memory Hierarchy Calculator AI Connector
Deterministic memory management for agentic memory tiers.
View
Suggestions
Agent Memory Tier Calculator AI Connector
Deterministic memory management engine for agentic memory hierarchies.
ViewPrompt Compression Efficiency Calculator AI Connector
Evaluate the performance, cost-effectiveness, and quality impact of prompt compression techniques.
ViewBuilder Iteration Learning Rate AI Connector
Analyzes learning velocity and execution efficiency in iterative development cycles.
ViewEdge Latency Simulator AI Connector
Estimates network latency for edge-computing deployment scenarios using geographic distance heuristics.
View
