Skip to content
#

g-eval

Here are 10 public repositories matching this topic...

Language: All
Filter by language

This repository provides a solution for generating detailed and thoughtful questions based on workout plans and evaluating their quality using the G-Eval metric. It is designed to assist fitness enthusiasts, trainers, and developers working with structured workout plans in improving the clarity, relevance, and usability of their questions.

  • Updated Dec 20, 2024
  • Python

DeepEval is an open-source LLM evaluation framework — built and maintained by Confident AI — for testing and benchmarking large language model applications. It is structured like Pytest but specialized for LLM systems, providing 40+ research-backed metrics (G-Eval, DAG, RAG metrics, agent metrics, multi-turn conversation metrics, multimodal…

  • Updated Aug 14, 2026

Improve this page

Add a description, image, and links to the g-eval topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the g-eval topic, visit your repo's landing page and select "manage topics."

Learn more