#Llm MCP Servers
Discover 21 MCP servers tagged with Llm on the Vinkius App Catalog.
OpenAI MCP
Use GPT-4o, DALL-E 3, embeddings, fine-tuning, and moderation as tools inside your AI agent workflows.
Anthropic MCP
Access Claude models via Anthropic API. Send messages, count tokens, manage batches and discover models from any AI agent.
Mistral AI (Frontier LLMs & Embeddings) MCP
Manage AI inference via Mistral. Execute chat completions, generate RAG embeddings, and audit frontier models.
NVIDIA AI MCP
Access LLMs, embeddings, code generation, and reasoning via NVIDIA API Catalog.
Cohere (AI Platform) MCP
Power enterprise AI via Cohere. Generate text, perform chat completions, reorder documents, and manage embeddings directly from any AI agent.
Cohere MCP
Access Cohere AI models via API. Chat with Command models, generate embeddings, rerank documents and tokenize text from any AI agent.
Mistral AI MCP
Access Mistral AI models via API. Chat with Claude alternatives, generate embeddings, moderate content and manage batch jobs from any AI agent.
DeepSeek MCP
Access powerful open-weight language models for reasoning, code generation, and complex problem solving at competitive cost.
Together AI MCP
Access 100+ open-source models for chat, image generation, and fine-tuning. Power your AI agents with Llama 3.3, Flux, and more.
NewOllama MCP
Run LLM models via Ollama cloud API. Generate completions, chat with multimodal models, create embeddings, and inspect model details from any AI agent.
NewZ.AI MCP
Access the full Z.AI platform from any AI agent. Chat completions with GLM models, image and video generation, audio transcription, OCR, web search, and agent tools.
Together AI MCP
Generate code, evaluate embeddings, and deploy open-source LLMs instantly from your local agent via Together AI's infrastructure.
Gradient AI (LLM API & Finetuning) MCP
Access powerful LLMs, fine-tune models on your own data, and generate embeddings directly through your AI agent.
Writer (AI Enterprise LLM) MCP
Access Writer's enterprise-grade LLMs and Knowledge Graph capabilities to generate content, manage files, and query RAG-based data.
Forefront MCP
Access Forefront AI models directly from your agent. Generate chat completions, manage fine-tuning jobs, and collect LLM outputs with pipelines.
NewGPU Inference Memory Calculator MCP
Estimate GPU VRAM requirements for LLM inference based on model parameters, precision, and batch size.
NewLLM API Cost Calculator MCP
Estimate and compare the financial impact of LLM usage across different providers.
NewLLM Context Window Budgeter MCP
Monitor and predict LLM context window exhaustion with precision token forecasting.
NewPrompt Injection Pattern Scanner MCP
Scans user-supplied text for structural patterns associated with prompt-injection attempts.
NewRAG Chunk Size Optimizer MCP
Evaluate RAG chunking strategies by calculating segmentation metrics, embedding costs, and context viability.
NewLLM Fine-Tuning Dataset Validator MCP
Verify structural integrity, token distribution, and training costs of JSONL datasets.