ClaudeChatGPTPerplexityGeminiMicrosoft CopilotRaycastMeta AIGrokZ.aiQwenKimi
DeepSeekMistralCursorVS CodeWindsurfJetBrainsClineLovableVercel AI SDKLangChain

Use NVIDIA API Catalog with your AI.

Connect your account once and let the AI you already use work with it, without building another integration. Cloud Engine proxy running native foundational completions natively utilizing active Nemotron and Llama3 architectures.

Included with plan

Ask AI about this Connector

Developed, maintained, and hosted by Vinkius.

MCP VERIFIED · PRODUCTION READY · VINKIUS GUARANTEED

Waiting for input…

Works with modern AI clients that support MCP, including ChatGPT, Claude, Cursor, and more.

ChatGPTClaudeCursorPerplexityGeminiMicrosoft CopilotRaycastMeta AI

Complete set · 8 capabilities

The complete NVIDIA API Catalog capability set.

These are the exact actions your AI can choose when you ask it to work with NVIDIA API Catalog.

Capability set01 / 02

01-04

4 capabilities in this set.

Part of 8 available through NVIDIA API Catalog.

  1. 01

    Nvidia list foundation models

    Dumps the strict array specifying explicit LLM matrix paths accessible securely natively

  2. 02

    Nvidia list lora adapters

    Evaluate explicit matrices tracking fine-tuned overrides isolating logical constraints dynamically

  3. 03

    Nvidia vision inference

    G. Llama-Vision natively). Invoke strictly multimodal abilities capturing diagnostic constraints returning inference on graphical data

  4. 04

    Nvidia chat completion

    Trigger direct NLP inference matrices directly evaluating queries over hosted LLMs

Capability set02 / 02

05-08

4 capabilities in this set.

Part of 8 available through NVIDIA API Catalog.

  1. 05

    Nvidia check token quota

    Poll safely dynamic credit and explicit constraint execution limits bounding inference execution

  2. 06

    Nvidia generate embeddings

    Pass parameters safely mapping explicit unstructured vectors directly using specific Embedding arrays

  3. 07

    Nvidia get cloud status

    Ping explicitly the core hosted NVIDIA matrix tracing inference endpoints evaluating latencies securely

  4. 08

    Nvidia summarize content

    Standard natively configured logical execution executing predefined abstract compression matrices smoothly

Observed, not estimated

826ms average. Fast in production.

NVIDIA API Catalog is checked daily against the live service.

Daily averagePeak 1029ms
Aug 20Today
Fastest day
677ms
Slowest day
1029ms
14-day trend
Slowing+15%

Connect your client

One URL. Every client.

Activate the Connector, copy your link, and paste it into the client you already use. 8 capabilities arrive ready to run.

Preview access · not provider authentication

The vk_preview_* token belongs to Vinkius preview infrastructure. It lets Claude discover and display the capabilities of NVIDIA API Catalog, so you can see the experience inside your AI.

It does not authenticate your account with NVIDIA API Catalog. Actions requiring credentials or live account data may not run until you activate the Connector and authorize the service.

NVIDIA API Catalog Connector

You're all set. Choose your MCP client and follow the setup instructions.

Connector linkhttps://edge.vinkius.com/vk_preview_LgfbZayzkE3ZBfq9HWv3rgeK69AXP0h3anUpakUy/mcp

Claude Desktop

Follow the steps below to connect in seconds.

  1. 1In Claude Desktop, open Settings → Connectors.
  2. 2Click “Add custom connector” and paste the connector link above as the remote MCP server URL.
  3. 3Click Add and start a new chat — NVIDIA API Catalog capabilities are ready to use.
Configuration · claude_desktop_config.jsonCopy
{
  "mcpServers": {
    "nvidia-api-catalog-mcp": {
      "url": "https://edge.vinkius.com/vk_preview_LgfbZayzkE3ZBfq9HWv3rgeK69AXP0h3anUpakUy/mcp"
    }
  }
}
  • Claude
  • ChatGPT
  • Cursor
  • VS Code
  • Windsurf
  • Claude Code
  • JetBrains
  • Cline

Step-by-step instructions for each client are in the guide. How to connect

FAQ

Questions NVIDIA API Catalog owners ask.

  • 01

    Can I explicitly route specific embedding vectors natively using the NVIDIA integration matrix?

    Yes! Utilize generate_embeddings providing explicit logic extracting arrays natively isolating endpoints safely.

  • 02

    How do I explicitly explore active LLMs natively hosted inside the NVIDIA catalog bounds?

    Target explicit matrices natively calling list_foundation_models returning catalog endpoints safely explicitly mapping bounds secure natively.

  • 03

    Does this require local Docker execution mapping explicitly NVIDIA parameters transparently?

    No, this explicitly pings the hosted Cloud API. For local Docker metrics natively, switch to nvidia-nim-mcp enforcing natively local boundaries.