eval-view

by hidai25

Proof your AI agent still works. Regression testing with golden baselines, tool-call diffing, and output drift detection. MCP server + Claude Code skills. LangGraph, CrewAI, Anthropic, OpenAI.

โ˜… 47 stars6 forksActivePython
79
Good

๐Ÿ“Š Score Breakdown

๐Ÿ›ก๏ธSecurity30%
5.0/5
โšกUtility30%
2.0/5
๐Ÿ”„Maintenance25%
5.0/5
๐Ÿ’ŽUniqueness15%
4.0/5

Overall = Security (30%) + Utility (30%) + Maintenance (25%) + Uniqueness (15%). Full methodology โ†’

โ„น๏ธ Details

Ecosystem

Claude Skill

Language

Python

Pricing

Free

License

Apache-2.0

Status

Active

๐Ÿ“ˆ GitHub Signals

47

Stars

6

Forks

0

Commits (30d)

12

Open Issues

Last commit: 4 months ago

agentagent-benchmarkagent-evaluationagentic-aiai-agentsanthropiccrewaicrewai-toolsevaluationlangchainlanggraphlanggraph-pythonllmllmopsmlopsopenai-assistantspytesttestingtools

๐Ÿ… Show your score

Scored 79/100 for security, utility and maintenance. Add the badge to your README or site to show it, verified by an independent directory.

eval-view scored 79/100 on SkillsIndex
Markdown
[![eval-view scored 79/100 on SkillsIndex](https://skillsindex.dev/api/badge/hidai25-eval-view)](https://skillsindex.dev/tools/hidai25-eval-view/)
HTML
<a href="https://skillsindex.dev/tools/hidai25-eval-view/"><img src="https://skillsindex.dev/api/badge/hidai25-eval-view" alt="eval-view scored 79/100 on SkillsIndex" height="20"></a>

Know before you install ๐Ÿ“ฌ

We score every tool 0-100 for security, maintenance and utility. Get the weekly shortlist of the highest-scored, vetted tools, plus an alert when a package you rely on goes stale. Free.

Data last verified: 4 months ago. See something wrong? Report it โ†’