claim
KGHaluBench incorporates a difficulty-scaled accuracy metric to ensure a fair and consistent benchmark, accounting for variations in difficulty caused by dynamic question generation.
Authors
Sources
- A Knowledge Graph-Based Hallucination Benchmark for Evaluating ... arxiv.org via serper
Referenced by nodes (1)
- KGHaluBench concept