deepeval-evaluation-harness MCP Server
sunilp303/deepeval-evaluation-harness
Pluggable DeepEval scaffold for RAG, agents, and LLM apps across Anthropic, Bedrock, Azure OpenAI, and Vertex. Ships traceability, test synthesis, safety/PII gating, multi-turn conversation eval, agentic tool-use scoring, JSON validation, judge benchmarks, hyperparameter sweeps, and pytest CI — one Makefile target per feature.
claude mcp add agentrank -- npx -y agentrank-mcp-server Overview
sunilp303/deepeval-evaluation-harness is a Python agent tool licensed under MIT. Pluggable DeepEval scaffold for RAG, agents, and LLM apps across Anthropic, Bedrock, Azure OpenAI, and Vertex. Ships traceability, test synthesis, safety/PII gating, multi-turn conversation eval, agentic tool-use scoring, JSON validation, judge benchmarks, hyperparameter sweeps, and pytest CI — one Makefile target per feature.
Ranked #21951 out of 24850 indexed tools.
Ecosystem
Score Breakdown
1 stars → early stage
Last commit 3mo ago → stale
No issues filed → no history to score
0 contributors → solo project
No dependents → no downstream usage
Weights: Freshness 20% · Issue Health 20% · Dependents 22% · Stars 10% · Contributors 8% · How we score →
How to Improve
Matched Queries
Get the weekly AgentRank digest
Top movers, new tools, ecosystem insights — straight to your inbox.