Category
Llm Inference
Llm Inference MCP Servers
Browse 2 Llm Inference MCP servers on the Vinkius MCP Catalog. Enterprise-grade connectors, operational in seconds.
New
Ollama MCP
12 tools
Run LLM models via Ollama cloud API. Generate completions, chat with multimodal models, create embeddings, and inspect model details from any AI agent.
NewGPU Inference Memory Calculator MCP
4 tools
Estimate GPU VRAM requirements for LLM inference based on model parameters, precision, and batch size.