ClaudeChatGPTPerplexityGeminiMicrosoft CopilotRaycastMeta AIGrokZ.aiQwenKimi
DeepSeekMistralCursorVS CodeWindsurfJetBrainsClineLovableVercel AI SDKLangChain

Use NVIDIA with your AI.

Connect your account once and let the AI you already use work with it, without building another integration. Transcribe speech, generate voices, translate audio, and clone voices via NVIDIA Audio APIs.

Included with plan

Ask AI about this Connector

Developed, maintained, and hosted by Vinkius.

MCP VERIFIED · PRODUCTION READY · VINKIUS GUARANTEED

Waiting for input…

Works with modern AI clients that support MCP, including ChatGPT, Claude, Cursor, and more.

ChatGPTClaudeCursorPerplexityGeminiMicrosoft CopilotRaycastMeta AI

Complete set · 10 capabilities

The complete NVIDIA capability set.

These are the exact actions your AI can choose when you ask it to work with NVIDIA.

Capability set01 / 03

01-04

4 capabilities in this set.

Part of 10 available through NVIDIA.

  1. 01

    List audio models

    List available audio models on NVIDIA API Catalog

  2. 02

    Punctuate text

    Add punctuation and capitalization to raw text

  3. 03

    Speaker diarization

    Identify different speakers in an audio file

  4. 04

    Text to speech

    Optional voice parameter for different voices. Convert text to natural-sounding speech

Capability set02 / 03

05-07

3 capabilities in this set.

Part of 10 available through NVIDIA.

  1. 05

    Audio translation

    Provide target language. Translate spoken audio to another language

  2. 06

    Cancel noise

    Remove background noise from audio

  3. 07

    Classify audio

    ) with confidence scores. Classify the type of sound in an audio file

Capability set03 / 03

08-10

3 capabilities in this set.

Part of 10 available through NVIDIA.

  1. 08

    Clone voice

    Clone a voice from a reference audio and generate speech

  2. 09

    Speech to text

    Supports multiple languages. Provide a public audio URL (MP3, WAV, etc). Transcribe speech from audio to text (Whisper-style)

  3. 10

    Summarize audio

    Summarize an audio transcript

Observed, not estimated

872ms average. Fast in production.

NVIDIA is checked daily against the live service.

Daily averagePeak 1122ms
Aug 20Today
Fastest day
739ms
Slowest day
1122ms
14-day trend
Slowing+25%

Connect your client

One URL. Every client.

Activate the Connector, copy your link, and paste it into the client you already use. 10 capabilities arrive ready to run.

Preview access · not provider authentication

The vk_preview_* token belongs to Vinkius preview infrastructure. It lets Claude discover and display the capabilities of NVIDIA, so you can see the experience inside your AI.

It does not authenticate your account with NVIDIA. Actions requiring credentials or live account data may not run until you activate the Connector and authorize the service.

NVIDIA Connector

You're all set. Choose your MCP client and follow the setup instructions.

Connector linkhttps://edge.vinkius.com/vk_preview_wisSzrtJnZMCHmz1iKrZAGZQEphbj0twPYlJjzMT/mcp

Claude Desktop

Follow the steps below to connect in seconds.

  1. 1In Claude Desktop, open Settings → Connectors.
  2. 2Click “Add custom connector” and paste the connector link above as the remote MCP server URL.
  3. 3Click Add and start a new chat — NVIDIA capabilities are ready to use.
Configuration · claude_desktop_config.jsonCopy
{
  "mcpServers": {
    "nvidia-audio-mcp": {
      "url": "https://edge.vinkius.com/vk_preview_wisSzrtJnZMCHmz1iKrZAGZQEphbj0twPYlJjzMT/mcp"
    }
  }
}
  • Claude
  • ChatGPT
  • Cursor
  • VS Code
  • Windsurf
  • Claude Code
  • JetBrains
  • Cline

Step-by-step instructions for each client are in the guide. How to connect

FAQ

Questions NVIDIA owners ask.

  • 01

    What languages are supported for transcription?

    Parakeel models support 50+ languages including English, Portuguese, Spanish, French, German, Mandarin, Japanese, and many more. Specify the language for best results.

  • 02

    Can I clone a specific voice?

    Yes! Use the clone_voice capability with a reference audio sample (a few seconds is enough) and the text you want the cloned voice to speak.

  • 03

    What is speaker diarization?

    Speaker diarization identifies 'who spoke when' in an audio recording. It segments the audio by speaker and returns timestamps for each speaker's turns.

  • 04

    What audio formats are supported?

    The API supports WAV, MP3, FLAC, OGG, and most common audio formats. For best transcription accuracy, use high-quality WAV or FLAC files at 16kHz or higher sample rate.