claim
The calibration dataset for hallucination evaluation captures the entity ID, entity statistics, entity type, question relation types, relation scores, and the overall question score for each question.
Authors
Sources
- A Knowledge Graph-Based Hallucination Benchmark for Evaluating ... arxiv.org via serper
Referenced by nodes (1)
- Hallucination Evaluation Model concept