mcp-ollama
About
MCP server wrapping local Ollama models for Claude Code delegation
Details
- License
- Apache-2.0
Explore
- Nine tools for generation, summarisation, code, diffs, and more
- Runs entirely over stdio – no network call‑outs by the server
- Override model per tool call or via OLLAMA_MODEL environment variable
- Ships no model weights, no telemetry, no analytics
- Stateless between calls; every request stays on the host running Ollama
- Supports Docker deployment with automatic host‑network routing
Setting up with Highlight
This MCP is not yet compatible with Highlight’s one-click setup. However, you can still use it with Highlight by following these steps:
- Download and install Highlight from highlightai.com/download
- Navigate to the plugins tab and select "Add Custom Plugin"
-
Configure the plugin with the settings below
Plugin Name
mcp-ollamaCommand (node, npx, python, etc.)Please refer to the README for specific instructions on how to obtain API keys or other required environment variables.
- Enable "Start Automatically" if you want the plugin to start when Highlight launches
From the repository
npm install -g @truealter/mcp-ollama
Or invoke directly without installing:
npx @truealter/mcp-ollama
You also need a running Ollama instance with at least one model pulled:
bash
claude mcp add --transport stdio ollama -- node /absolute/path/to/mcp-ollama/dist/index.js
Or in ~/.claude/settings.json:
json{
"mcpServers": {
"ollama": {
"transport": "stdio",
"command": "node",
"args": ["/absolute/path/to/mcp-ollama/dist/index.js"],
"env": {
"OLLAMA_HOST": "http://localhost:11434",
"OLLAMA_MODEL": "hermes3:8b"
}
}
}
}
``
| Variable | Default | Purpose |
|----------------|-----------------------------|---------------------------------------------------|
|
OLLAMA_HOST | http://localhost:11434 | Ollama HTTP endpoint |
| OLLAMA_MODEL | hermes3:8b | Default model when a tool call omits model |
Any tool call may override
model explicitly - the env default only applies when unset. local_code tends to work better with a code-specialised model passed per-call, while local_summarize and local_draft` are fine on the default.local_generate
General-purpose generation with system + user prompt
local_summarize
Summarise a blob of text
local_analyze
Analyse text against a specific question
local_draft
Draft content in a given style
local_code
Code tasks: docstring / test / explain / review / types / refactor-suggest
local_diff
Diff-driven tasks: commit-message / pr-description / changelog / summary / impact
local_transform
Mechanical code transformations
local_models
List models available on the local Ollama host
local_pull
Pull a model onto the local Ollama host
| Tool | Purpose |
|-------------------|--------------------------------------------------------------------------|
| local_generate | General-purpose generation with system + user prompt |
| local_summarize | Summarise a blob of text |
| local_analyze | Analyse text against a specific question |
| local_draft | Draft content in a given style |
| local_code | Code tasks: docstring / test / explain / review / types / refactor-suggest |
| local_diff | Diff-driven tasks: commit-message / pr-description / changelog / summary / impact |
| local_transform | Mechanical code transformations |
| local_models | List models available on the local Ollama host |
| local_pull | Pull a model onto the local Ollama host |
Full tool schemas are exposed over MCP introspection - any MCP-aware client will enumerate them automatically.
Claude Desktop / Cursor
Paste into your MCP client config file to install this server.
{
"mcpServers": {
"mcp-ollama": {
"ollama": {
"command": "node",
"args": [
"/absolute/path/to/mcp-ollama/dist/index.js"
],
"env": {
"OLLAMA_HOST": "http://localhost:11434",
"OLLAMA_MODEL": "hermes3:8b"
}
}
}
}
}
McpServers
{
"ollama": {
"command": "node",
"args": [
"/absolute/path/to/mcp-ollama/dist/index.js"
],
"env": {
"OLLAMA_HOST": "http://localhost:11434",
"OLLAMA_MODEL": "hermes3:8b"
}
}
}
MCP server wrapping local Ollama models for offload from API-priced orchestrators.
Exposes nine tools that pass work to a local model (text generation, summarisation, code tasks, mechanical transforms, commit/PR/changelog drafting). The orchestrator decides what to route locally; this server does the routing.
- Transport: stdio
- Runtime: Node 18+
- Default model: hermes3:8b (override via OLLAMA_MODEL)
- Ollama host: http://localhost:11434 (override via OLLAMA_HOST)
- Ships no model weights, no cloud call-outs, no telemetry. Every request stays on the host where Ollama is running.
- License: Apache-2.0
Why
Orchestrators priced by the token (Claude Code, Cursor, the Anthropic API, Cline, Aider) pay for every classification, every docstring, every commit message. Most of that work doesn't need a frontier model. Routed to Ollama on the same machine, the same work is free and faster. mcp-ollama is the routing surface.
The orchestrating model decides what to route where. This server is plumbing - it does not try to be clever about task classification. Pick the right tool, pass the text, get a result back.
Install
Install (npm)
npm install -g @truealter/mcp-ollama
Or invoke directly without installing:
npx @truealter/mcp-ollama
You also need a running Ollama instance with at least one model pulled:
```bash
Sign in to leave a review
Use Google, GitHub, or an email account so ratings stay tied to real people.
No reviews posted yet.



