N/A
SaiArja/LLM Evaluation Harness MCP Server
Evaluates RAG outputs on faithfulness, answer relevancy, and context precision using an LLM-as-a-Judge backend. Exposes tools for running evaluations, scoring individual samples, and checking thresholds, enabling CI gating and on-demand assessment via MCP.
Scan Scheduled
This agent is queued for security scanning. It will be graded in the next scan batch.
What We Know
- URL https://github.com/SaiArja/llm-eval-mcp
- Framework mcp
- Sources glama
- First Seen Jul 13, 2026
- Repository github.com/SaiArja/llm-eval-mcp
Browse more:
Search all agents
Ecosystem Report