eval-view

by hidai25 · MCP 服务器 · ★ 114

About eval-view

The open-source behavior regression gate for AI agents. Think Playwright, but for tool-calling and multi-turn AI agents. Your agent can still return and be wrong. A model or provider update can change tool choice, skip a clarification, or degrade output quality without changing your code or breaking

agent-benchmarkagent-evaluationagentic-aiai-agentsanthropicautogenclicrewaievaluationlangchain-agent

Quick Facts

Stars114
Forks20
LanguagePython
CategoryMCP 服务器
LicenseApache-2.0
Quality Score54.178/100
Open Issues6
Last Updated2026-06-15
Created2025-11-17
Platformscli, mcp, python
Est. Tokens~24k

More MCP 服务器 Tools

Explore other popular mcp 服务器 tools:

View all MCP 服务器 tools →

Popular Python Agent Tools

Frequently Asked Questions

What is eval-view?

eval-view is Regression testing for AI agents. Snapshot behavior,diff tool calls,catch regressions in CI. Works with LangGraph, CrewAI, OpenAI, Anthropic.. It is categorized as a MCP 服务器 with 114 GitHub stars.

What programming language is eval-view written in?

eval-view is primarily written in Python. It covers topics such as agent-benchmark, agent-evaluation, agentic-ai.

How do I install or use eval-view?

You can find installation instructions and usage details in the eval-view GitHub repository at github.com/hidai25/eval-view. The project has 114 stars and 20 forks, indicating an active community.

What license does eval-view use?

eval-view is released under the Apache-2.0 license, making it free to use and modify according to the license terms.

View on GitHub → Browse MCP 服务器 tools