Use NVIDIA NIM with your AI.
Connect your account once and let the AI you already use work with it, without building another integration. MLOps proxy unifying explicitly local hardware limits extracting telemetry across active NVIDIA AI containers.
Developed, maintained, and hosted by Vinkius.
MCP VERIFIED · PRODUCTION READY · VINKIUS GUARANTEED
Waiting for input…
Works with modern AI clients that support MCP, including ChatGPT, Claude, Cursor, and more.
Complete set · 8 capabilities
The complete NVIDIA NIM capability set.
These are the exact actions your AI can choose when you ask it to work with NVIDIA NIM.
01-04
4 capabilities in this set.
Part of 8 available through NVIDIA NIM.
- 01
Nim check health live
Execute liveness probes natively evaluating if the physical host container orchestrator is responsive
- 02
Nim check health ready
Detect if the GPU inference layers have successfully loaded the explicitly configured model artifacts natively
- 03
Nim get container logs
Fetch explicit execution parameters catching native stdout proxies bound cleanly to the orchestrator layer securely
- 04
Nim get gpu status
Parse explicit GPU topological limits mapped onto the NIM proxy securely formatting active hardware memory variables cleanly
05-08
4 capabilities in this set.
Part of 8 available through NVIDIA NIM.
- 05
Nim get metadata
Pull logical engine execution metrics mapping exactly the loaded foundational configuration bounds natively secure
- 06
Nim get metrics
Extract Prometheus hardware scaling metrics explicitly from the NIM orchestrator natively
- 07
Nim list models
Dump explicit active LLMs securely allocating inference targets over the logical backend array cleanly
- 08
Nim scale replicas
Dynamically orchestrate bounds adjusting native hardware replication proxy assignments scaling execution layers
Observed, not estimated
803ms average. Fast in production.
NVIDIA NIM is checked daily against the live service.
- Fastest day
- 638ms
- Slowest day
- 1125ms
- 14-day trend
- Slowing+15%
Connect your client
One URL. Every client.
Activate the Connector, copy your link, and paste it into the client you already use. 8 capabilities arrive ready to run.
Preview access · not provider authentication
The vk_preview_* token belongs to Vinkius preview infrastructure. It lets Claude discover and display the capabilities of NVIDIA NIM, so you can see the experience inside your AI.
It does not authenticate your account with NVIDIA NIM. Actions requiring credentials or live account data may not run until you activate the Connector and authorize the service.
NVIDIA NIM Connector
You're all set. Choose your MCP client and follow the setup instructions.
https://edge.vinkius.com/vk_preview_70BV0qtRSn6oYi65LbVY4KDoGOvWiulNTIr6cS4G/mcpClaude Desktop
Follow the steps below to connect in seconds.
- 1In Claude Desktop, open Settings → Connectors.
- 2Click “Add custom connector” and paste the connector link above as the remote MCP server URL.
- 3Click Add and start a new chat — NVIDIA NIM capabilities are ready to use.
{
"mcpServers": {
"nvidia-nim-mcp": {
"url": "https://edge.vinkius.com/vk_preview_70BV0qtRSn6oYi65LbVY4KDoGOvWiulNTIr6cS4G/mcp"
}
}
}
Claude
ChatGPT
Cursor
VS Code
Windsurf
Claude Code
JetBrains
Cline
Step-by-step instructions for each client are in the guide. How to connect
FAQ
Questions NVIDIA NIM owners ask.
- 01
Can I explicitly track GPU hardware analytics natively using the NIM MCP integration?
Yes! Utilize get_metrics exposing Prometheus-compatible proxy limits tracking explicit hardware latencies easily natively securely.
- 02
How do I explicitly evaluate if my container instances mapped properly loaded native Foundation Models?
Target UUID probes natively mapped executing check_health_ready verifying bounds catching limits generating exact readiness states cleanly.
- 03
Does this call inference proxies executing completions bounds mapped dynamically?
No, this is infrastructure proxy bounding explicitly container node management. Utilize nvidia-catalog-mcp enforcing natively hosted inference bounds efficiently.
Explore
More in Industry Titans
NVIDIA API Catalog AI Connector
Cloud Engine proxy running native foundational completions natively utilizing active Nemotron and Llama3 archi
ViewLangfuse (LLM Tracing & Evals) AI Connector
Monitor LLM apps via Langfuse — track traces, manage prompt templates, and audit evaluation scores.
ViewLiteLLM (LLM Proxy & Spend Tracking) AI Connector
Manage your LLM gateway via LiteLLM — generate API keys, track spending, and orchestrate model fallback paths.
ViewFireworks AI AI Connector
Empower LLM applications via Fireworks AI — perform ultra-fast chat completions, generate embeddings and image
View
Suggestions
NMKR Cardano AI Connector
Manage Cardano NFT projects via NMKR Studio — track assets, minting coupons, and payout wallets directly from
ViewNorthflank (Developer Cloud & Orchestration) AI Connector
Manage cloud infrastructure via Northflank — deploy microservices, trigger CI builds, and audit background job
ViewNetlify AI Connector
Modern web development platform — manage sites, deploys, and forms via AI.
ViewCTO Architect Prover AI Connector
An AI proposed Kubernetes for 50 users, says 'use HTTPS' as a security strategy, and plans database migrations
View
