> ## Documentation Index
> Fetch the complete documentation index at: https://docs.trusys.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# Functional Evaluation

TRU EVAL enables structured, scalable evaluation of AI models across custom or standardized metrics.

<img src="https://mintcdn.com/trusys/SzEVAMEivKYXTYeV/images/trueval-i.png?fit=max&auto=format&n=SzEVAMEivKYXTYeV&q=85&s=9735177417e00d82cb548f324c8cb4f4" alt="Descriptive alt text" noZoom height="200" data-path="images/trueval-i.png" />

### Key Capabilities

* Run prompt-based evaluations against custom datasets or shared benchmarks
* Compare multiple models (e.g., GPT-4, Claude, custom LLMs)
* Supports both **reference-based** and **referenceless evaluation** (also known as “LLM-as-a-judge”)
* Visualize performance across key capabilities such as factuality, reasoning, helpfulness
* Track regression or improvement in model behavior over time

### Use Cases

* Score open-source vs proprietary models
* Tailor evaluation templates to organizational needs
* Measure model alignment to domain-specific requirements
