Use Predibase with your AI.
Connect your account once and let the AI you already use work with it, without building another integration. Deploy and query fine-tuned LLMs via Predibase. run inference, classify text, and monitor deployment metrics directly from your AI agent.
Developed, maintained, and hosted by Vinkius.
MCP VERIFIED · PRODUCTION READY · VINKIUS GUARANTEED
Waiting for input…
Works with modern AI clients that support MCP, including ChatGPT, Claude, Cursor, and more.
Complete set · 7 capabilities
The complete Predibase capability set.
These are the exact actions your AI can choose when you ask it to work with Predibase.
01-04
4 capabilities in this set.
Part of 7 available through Predibase.
- 01
Classify
Batch classification for one or more inputs
- 02
Get health
Check health status of the inference endpoint
- 03
Get metrics
Get Prometheus metrics for the deployment
- 04
Chat completion
Create a chat completion (OpenAI compatible)
05-07
3 capabilities in this set.
Part of 7 available through Predibase.
- 05
Completion
Create a completion (OpenAI compatible)
- 06
Generate text
Generate text using a deployed LLM
- 07
Get info
Get inference endpoint metadata
Observed, not estimated
900ms average. Fast in production.
Predibase is checked daily against the live service.
- Fastest day
- 721ms
- Slowest day
- 1244ms
- 14-day trend
- Slowing+24%
Connect your client
One URL. Every client.
Activate the Connector, copy your link, and paste it into the client you already use. 7 capabilities arrive ready to run.
Preview access · not provider authentication
The vk_preview_* token belongs to Vinkius preview infrastructure. It lets Claude discover and display the capabilities of Predibase, so you can see the experience inside your AI.
It does not authenticate your account with Predibase. Actions requiring credentials or live account data may not run until you activate the Connector and authorize the service.
Predibase Connector
You're all set. Choose your MCP client and follow the setup instructions.
https://edge.vinkius.com/vk_preview_anUwoOMv2E6QQVudvBXsDytO30kiB1fpaXuzKUbp/mcpClaude Desktop
Follow the steps below to connect in seconds.
- 1In Claude Desktop, open Settings → Connectors.
- 2Click “Add custom connector” and paste the connector link above as the remote MCP server URL.
- 3Click Add and start a new chat — Predibase capabilities are ready to use.
{
"mcpServers": {
"predibase-llm-serving-finetuning-mcp": {
"url": "https://edge.vinkius.com/vk_preview_anUwoOMv2E6QQVudvBXsDytO30kiB1fpaXuzKUbp/mcp"
}
}
}
Claude
ChatGPT
Cursor
VS Code
Windsurf
Claude Code
JetBrains
Cline
Step-by-step instructions for each client are in the guide. How to connect
FAQ
Questions Predibase owners ask.
- 01
Can I use my fine-tuned adapters with this server?
Yes. When using the generate_text capability, you can provide an adapter_id to apply your specific fine-tuned LoRA adapter to the base model deployment.
- 02
How do I monitor the performance of my Predibase deployment?
Use the get_metrics capability to scrape Prometheus-formatted metrics or get_info to retrieve metadata like model ID and device type.
- 03
Does this support structured JSON responses?
Absolutely. The generate_text capability includes a schema parameter that allows you to pass a JSON schema to ensure the model output follows a specific structure.
Explore
More in Developer Tools
Cohere (AI Platform) AI Connector
Power enterprise AI via Cohere — generate text, perform chat completions, reorder documents, and manage embedd
ViewLLM Fine-Tuning Dataset Validator AI Connector
Verify structural integrity, token distribution, and training costs of JSONL datasets.
ViewAbacus AI (Enterprise AI Cloud) AI Connector
Manage the full machine learning lifecycle via Abacus AI — create projects, train models, and deploy real-time
ViewNyckel ML AI Connector
Classify data and perform semantic search via Nyckel — track ML functions, samples, and labels directly from y
View
Suggestions
Cohere AI Connector
Access Cohere AI models via API — chat with Command models, generate embeddings, rerank documents and tokenize
ViewConfusion Matrix Engine AI Connector
Deterministically calculate True Positives, FP, Precision, Recall, F1-Score, and Accuracy local. Stop LLM hall
ViewTokenization Normalizer AI Connector
Resolves tokenization drift by normalizing text to match specific LLM tokenizer profiles.
ViewContext Window Optimizer AI Connector
Optimizes LLM context windows by selecting the most relevant and recent information within token limits.
View
