Fact — measurement — Knowledge Tree

The authors created a calibration dataset for hallucination evaluation by evaluating 34 models (ranging from 1B to 685B parameters) across 10 runs of 150 questions, generating 51,000 data points.

Authors

Person: Not available Organization: arXiv
A Knowledge Graph-Based Hallucination Benchmark for Evaluating ...

Sources

A Knowledge Graph-Based Hallucination Benchmark for Evaluating ... arxiv.org arXiv via serper

Referenced by nodes (1)

Hallucination Evaluation Model concept