Skip to content
appsgit

Evalview MCP

Evalview MCP is an MCP server that adds cloud and DevOps tools to AI assistants such as Claude Desktop, Claude Code and Cursor. Regression testing for AI agents. Golden baselines, CI/CD, LangGraph, CrewAI, OpenAI, Claude. It has 137 GitHub stars, is released under the Apache-2.0 license and runs locally with uvx evalview.

github.com/hidai25/eval-view (opens in a new tab)

Install Evalview MCP

Generated from the server's MCP registry entry. Replace your-value with your own values.

Claude Desktop

claude_desktop_config.json
{
  "mcpServers": {
    "evalview-mcp": {
      "command": "uvx",
      "args": [
        "evalview"
      ],
      "env": {
        "OPENAI_API_KEY": "your-value"
      }
    }
  }
}

Settings > Developer > Edit Config. macOS: ~/Library/Application Support/Claude/, Windows: %APPDATA%\Claude\. Restart Claude Desktop afterwards.

Claude Code

claude mcp add --env OPENAI_API_KEY=your-value --transport stdio evalview-mcp -- uvx evalview

Cursor

.cursor/mcp.json
{
  "mcpServers": {
    "evalview-mcp": {
      "type": "stdio",
      "command": "uvx",
      "args": [
        "evalview"
      ],
      "env": {
        "OPENAI_API_KEY": "your-value"
      }
    }
  }
}

Project file; use ~/.cursor/mcp.json to enable it in every project.

VS Code

.vscode/mcp.json
{
  "inputs": [
    {
      "type": "promptString",
      "id": "openai_api_key",
      "description": "OPENAI_API_KEY",
      "password": true
    }
  ],
  "servers": {
    "evalview-mcp": {
      "type": "stdio",
      "command": "uvx",
      "args": [
        "evalview"
      ],
      "env": {
        "OPENAI_API_KEY": "${input:openai_api_key}"
      }
    }
  }
}

Config formats checked against the official docs on Oct 7, 2026: modelcontextprotocol.io (opens in a new tab), code.claude.com (opens in a new tab), cursor.com (opens in a new tab), code.visualstudio.com (opens in a new tab).

Environment variables

Variables the server reads at startup.

NameRequiredDescription
OPENAI_API_KEYsecretNoOpenAI API key for LLM-as-judge output quality scoring. Optional — deterministic tool/sequence evaluation works without it.

About Evalview MCP

Snapshot testing for AI agents. Record what your agent does today. Get told when it silently changes. Your agent returns 200 and looks fine. But a model update, a provider change, or a one-line prompt edit just made it skip a clarification, call the wrong tool, or quietly drop output quality. Your tests still pass. Your users notice before you do.

  • agent-benchmark
  • agent-evaluation
  • ai-agents
  • crewai
  • evaluation
  • langgraph
  • openai-assistants
  • testing
  • pytest
  • anthropic

FAQ

Evalview MCP FAQ

Still curious? Email info@appsgit.com.

What is Evalview MCP?

Evalview MCP is an MCP server that adds cloud and DevOps tools to AI assistants such as Claude Desktop, Claude Code and Cursor. Regression testing for AI agents. Golden baselines, CI/CD, LangGraph, CrewAI, OpenAI, Claude. It has 137 GitHub stars, is released under the Apache-2.0 license and runs locally with uvx evalview. The source code is at github.com/hidai25/eval-view.

How do I install the Evalview MCP MCP server?

Add the command uvx evalview to your MCP client: put it in claude_desktop_config.json for Claude Desktop, run claude mcp add for Claude Code, or add it to .cursor/mcp.json (Cursor) or .vscode/mcp.json (VS Code). The snippets on this page are ready to paste.

Is Evalview MCP free?

The server is open source under the Apache-2.0 license, so running it is free. It does not declare any required API key.

Is Evalview MCP actively maintained?

The most recent commit was on Sep 5, 2026. The latest release is v0.8.1, published Jul 26, 2026. appsgit only lists MCP servers with a commit in the last six months and re-checks every server daily.