Skip to content
Vinkius

Replicate Connector for AI agents.

20 live capabilities

Run and manage machine learning models without setting up a server.

Live agent request Replicate / Connector

Waiting for input…

AI Agent

Why people use Replicate

Replicate for Machine Learning Model Deployment

This Connector removes that entire layer of friction. You just tell your agent what you want to do, and it handles the calls to Replicate. You get to stay in your flow, while the agent handles the model discovery, prediction execution, and status tracking. You get to focus on the results, not the rack.

  • Claude
  • ChatGPT
  • Gemini
  • Cursor
  • Visual Studio Code
  • Windsurf

What Vinkius changes

You get instant access to production-grade machine learning without managing any infrastructure.

Use it from Claude, ChatGPT, Cursor or another AI client you already have.

One account · 5,900+ Connectors

  1. Real-world use case 01

    Prompting for Art

    A designer wants a specific style of image.

  2. Real-world use case 02

    Fine-tuning a Niche Model

    A developer needs a model trained on specific company data.

  3. Real-world use case 03

    Scaling a Production App

    An engineer needs to serve model results to thousands of users.

Complete set · 20capabilities

The complete Replicate capability set.

These are the exact actions your AI can choose when you ask it to work with Replicate.

Capability set01 / 05

01—04

4 capabilities in this set.

Part of 20 available through Replicate.

  1. 01 Capability

    Cancel prediction

    Stop a prediction that is currently running. Use this if you notice a mistake in the input.

  2. 02 Capability

    Create deployment prediction

    Run a prediction using a dedicated deployment. This ensures faster, more reliable results for production.

  3. 03 Capability

    Create model

    Create a new model entry on the Replicate platform. Use this to organize your custom assets.

  4. 04 Capability

    Create prediction

    Trigger a model to run a specific task. It sends your inputs to the model and starts the process.

Capability set02 / 05

05—08

4 capabilities in this set.

Part of 20 available through Replicate.

  1. 05 Capability

    Create training

    Start a new training session to fine-tune a model. This is how you customize models on your own data.

  2. 06 Capability

    Delete model version

    Remove a specific version of a model. Use this to clean up old or broken versions.

  3. 07 Capability

    Get account

    See your account details and organization info. It's the quickest way to check your current status.

  4. 08 Capability

    Get collection

    Pull details for a specific model collection. This helps you see how models are grouped.

Capability set03 / 05

09—12

4 capabilities in this set.

Part of 20 available through Replicate.

  1. 09 Capability

    Get model version

    Get the details and OpenAPI schema for a specific version. This lets your agent understand the exact inputs needed.

  2. 10 Capability

    Get prediction

    Check the status and final output of a prediction. Use this to see if your generation finished.

  3. 11 Capability

    Get training

    Check the current status of a training job. This tells you if your fine-tuning is still running.

  4. 12 Capability

    Get webhook secret

    Retrieve the secret key for your webhooks. You'll need this to verify signatures from Replicate.

Capability set04 / 05

13—16

4 capabilities in this set.

Part of 20 available through Replicate.

  1. 13 Capability

    List collections

    See all curated collections of models. This is a good starting point for finding new capabilities.

  2. 14 Capability

    List hardware

    See all available hardware SKUs. Use this to understand the compute options available to you.

  3. 15 Capability

    List predictions

    See a history of your recent predictions. This is your go-to for looking back at previous runs.

  4. 16 Capability

    Search models

    Search for public models on Replicate. Use this to find the right capability for any ML task.

Capability set05 / 05

17—20

4 capabilities in this set.

Part of 20 available through Replicate.

  1. 17 Capability

    Update model

    Change the metadata for an existing model. This helps you keep your custom model info current.

  2. 18 Capability

    Create deployment

    Set up a private deployment with specific autoscaling rules. This is great for production environments.

  3. 19 Capability

    Get model

    Fetch specific details about a model. This is useful for checking configuration and metadata.

  4. 20 Capability

    List model versions

    See every version of a specific model. This helps you pick the right one for your task.

Set up in minutes

One URL. Then ask Replicate to work.

Claude and ChatGPT only need the Connector URL. Copy it once, add it in settings, and use Replicate from the conversation.

Choose your client

Live preview
Advanced clients IDE · CLI

Claude · Web + desktop

Official guide ↗

Connector URL · ready to paste

Streamable HTTP
https://edge.vinkius.com/vk_preview_ULeVbcRjoLCkIr9x93q2HML8VJd4AXmk4CWHHPso/mcp
  1. Step 01

    Open Connectors

    In Claude Web or Claude Desktop, open Settings and choose Connectors.

  2. Step 02

    Add the URL

    Choose Add custom connector, name it Replicate, and paste the URL above.

  3. Step 03

    Turn it on in chat

    Select +, open Connectors, and enable Replicate for the conversation.

Where the request belongs

Work Replicate can move forward.

Built around the request

This is for the developer who needs to run heavy ML models but doesn't want to spend a weekend configuring GPU drivers or managing a fleet of inference servers.

01

AI Engineer

Testing different model versions and parameters quickly without writing boilerplate API calls.

02

Creative Technologist

Generating high-quality media assets like images or audio directly within their design workflow.

03

Data Scientist

Monitoring long-running training sessions and checking prediction logs from a single interface.

Bring your own AI

Change the model, client or framework. Keep Replicate connected.

  • Claude
  • ChatGPT
  • Gemini
  • Cursor
  • VS Code
  • Windsurf
  • ZCode
  • Cline
  • Zed
  • Continue
  • Kiro
  • Roo Code
  • Zencoder
  • Goose
  • Void
  • Augment Code
  • Amp
  • Qodo
  • Tabnine
  • Pieces
  • Sourcegraph Cody
  • JetBrains
  • Warp
  • Amazon Q
  • Antigravity
  • BoltAI
  • Raycast
  • Jan
  • LM Studio
  • AnythingLLM
  • Open WebUI
  • Msty
  • Cherry Studio
  • LibreChat
  • TypingMind
  • Chorus
  • 5ire
  • n8n
  • LangChain
  • LlamaIndex
  • CrewAI
  • Vercel AI SDK

Before you connect

Questions about Replicate.

The practical details behind the request, access and result.

How does the Replicate MCP help with machine learning?

It lets your AI agent execute models directly. Instead of writing code to call an API, you just tell the agent what to generate or process, and it handles the communication with Replicate's infrastructure.

Can I use this Connector to run Stable Diffusion?

Yes, you can use search_models to find the best version and then run predictions to generate images directly from your chat.

How do I manage my custom models with the Replicate MCP?

You can use it to create, update, and delete models. Your agent can also pull specific versions and their schemas to ensure it's sending the right data.

Can my agent monitor my model training?

Absolutely. Use the get_training capability to check progress. Your agent can even give you updates as the training moves through different stages.

Is it possible to scale my model for production?

Yes, you can use create_deployment to set up a private deployment with autoscaling, which is perfect for handling high traffic.

How do I know what inputs a specific Replicate model needs?

The Connector can fetch the OpenAPI schema for any model version. This allows your agent to see exactly what parameters are required before it tries to run a prediction.

How can I check if my prediction has finished and see the output?

Use the get_prediction capability with your Prediction ID. It will return the current status (starting, processing, succeeded, or failed) along with the output URLs or data once completed.

Can I search for specific types of models like 'image-to-text'?

Yes! Use the search_models capability with your query. It will return a list of public models matching your terms, including their owners and descriptions.

Is it possible to stop a model that is taking too long to run?

Absolutely. Use the cancel_prediction capability with the target Prediction ID to immediately stop the execution and prevent further usage costs.

One connection away

Give your agent a direct line to Replicate.

Connect Replicate once. Keep it beside 5,900+ managed Connectors when the next task needs more.

Explore every Connector No credit card required · Free tier available