Best Ollama MCP Servers
Ollama is an open-source project that allows you to run large language models (LLMs) locally on your own hardware, providing a way to use AI capabilities privately without sending data to external services.
Why this server?
Leverages local Ollama LLM for offline diagnosis and remediation, avoiding external API calls.
AlicenseBqualityDmaintenanceAgentic data quality MCP server — runs structured validation rules against warehouses (DuckDB, BigQuery, Athena, Databricks, Postgres), diagnoses failures with LLM root cause analysis, and proposes SQL remediations. Full audit trail of every AI decision.64Apache 2.0Why this server?
Allows using local Ollama models as workers in MergeLoop, enabling on-device inference alongside other workers.

MergeLoopofficial
AlicenseAqualityCmaintenanceA host-agnostic model council that routes tasks across multiple AI workers (MCP, CLI, API) and returns one unified answer.13Apache 2.0Why this server?
Provides a gateway to route and load-balance requests across multiple Ollama instances, with session affinity and latency-based routing.
AlicenseAqualityAmaintenanceReal-time cluster health monitoring, pre-request NLMS latency prediction, and intelligent prompt routing across multi-instance LLM backends (vLLM, Ollama, SGLang, TGI).443Apache 2.0Why this server?
Integrates with Ollama for local AI-powered natural language to SQL query generation.
Why this server?
Uses Ollama as an embeddings backend to enable semantic and graph-seeded retrieval across vault notes.
AlicenseAqualityAmaintenanceModel-agnostic, agent-ready Obsidian MCP server with RBAC, SLSA provenance, and native search. 163 tools across 31 domains, multi-vault, pluggable embeddings.435AGPL 3.0Why this server?
Allows managing local Ollama models: discover installed and loaded models, pull/remove models, load/unload from memory, check hardware fit, and offload inference (completion and embedding) to local Ollama instances.

Local AI MCPofficial
AlicenseAqualityAmaintenanceUnified MCP server for managing local model runtimes (Ollama, LM Studio, etc.), enabling provider-agnostic discovery, lifecycle management, hardware-fit checks, and delegated inference.1630 npmCreative Commons Attribution Non Commercial No Derivatives 4.0 InternationalWhy this server?
Allows querying local Ollama models via HTTP.

MCP Rubber Duckofficial
AlicenseAqualityBmaintenanceAn MCP server that bridges multiple LLMs (OpenAI-compatible APIs and CLI coding agents) for collaborative debugging and diverse AI perspectives.12141 npm5MITWhy this server?
Provides a preset tunnel configuration for Ollama with authentication enabled on port 11434, enabling secure external access to local Ollama instances.
Why this server?
Provides tools for predicting LLM performance and optimizing local LLM inference using Ollama, including checking compatibility of specific models and getting recommendations.

Yamaru Hardware Probeofficial
AlicenseAqualityDmaintenanceExpert system hardware probe and performance diagnostic engine for AI, Gaming, and High-Performance workflows. Provides deep system insights such as real-time monitoring, thermal diagnostics, and LLM optimization.1128 npm7Apache 2.0