LLM Eval Search
Verified IntegrationClient Configuration
— Connect LLM Eval Search to Claude Desktop or Cursor in seconds{
"mcpServers": {
"llm-eval-search": {
"command": "npx",
"args": [
"-y",
"@modelcontextprotocol/server-llm-eval-search"
],
"env": {}
}
}
}~/Library/Application Support/Claude/claude_desktop_config.json (macOS) or %APPDATA%\Claude\claude_desktop_config.json (Windows).System Overview
Offers a production-grade framework for evaluating large language model answers using agentic orchestration, multi-metric analysis, and robust guardrails.
7/23/2026
Open Source
stdio / SSE RPC
Frequently Asked Questions
Architecture and operational details for LLM Eval Search
Acknowledging that no single metric is universally reliable, LLM Eval Search leverages multiple metrics to assess different facets of LLM answer quality. It flags metric disagreements for human review, providing a more comprehensive and robust evaluation.
Related MCP Servers
Browse all servers →Empower AI assistants to generate and convert documents, manage templates, and automate document workflows efficiently.
Provides a multi-tenant, AI-native Content Delivery Network deployable on Cloudflare, featuring sub-100ms TTFB, AI agent controllability, and comprehensive accessibility.
Deploy a Model Context Protocol server on Cloudflare Workers without requiring authentication.