Local LLM Delegation
Verified IntegrationClient Configuration
— Connect Local LLM Delegation to Claude Desktop or Cursor in seconds{
"mcpServers": {
"local-llm-delegation": {
"command": "npx",
"args": [
"-y",
"@modelcontextprotocol/server-local-llm-delegation"
],
"env": {}
}
}
}~/Library/Application Support/Claude/claude_desktop_config.json (macOS) or %APPDATA%\Claude\claude_desktop_config.json (Windows).System Overview
Intelligently offloads low-complexity AI agent tasks to local LLMs, preserving premium tokens for complex reasoning.
7/23/2026
Open Source
stdio / SSE RPC
Frequently Asked Questions
Architecture and operational details for Local LLM Delegation
It integrates seamlessly with LiteLLM, supporting a wide range of providers including Ollama, vLLM, LM Studio, Anthropic, and any OpenAI-compatible provider. For AI agent clients, it works with Gemini Code CLI and Claude Code.
Related MCP Servers
Browse all servers →Empower AI assistants to generate and convert documents, manage templates, and automate document workflows efficiently.
Provides a multi-tenant, AI-native Content Delivery Network deployable on Cloudflare, featuring sub-100ms TTFB, AI agent controllability, and comprehensive accessibility.
Deploy a Model Context Protocol server on Cloudflare Workers without requiring authentication.