Pay-per-call AI evaluation engine. Score LLM outputs and agent trajectories against benchmark rubrics. $0.005 per eval via x402 on Base.
-
Updated
Jul 17, 2026
Pay-per-call AI evaluation engine. Score LLM outputs and agent trajectories against benchmark rubrics. $0.005 per eval via x402 on Base.
🎯 现代化大语言模型标准自动化评测平台 (LLMEval) | 支持开源/商业模型 · 五维能力雷达图 · Bad Case 逐题审查 · SQLite 历史持久化 · Gradio Web & CLI
Add a description, image, and links to the eval-engine topic page so that developers can more easily learn about it.
To associate your repository with the eval-engine topic, visit your repo's landing page and select "manage topics."