Fact — measurement — Knowledge Tree

Applying difficulty-based weighting to the benchmark decreases the mean accuracy by 0.05%, from 45.30% to 45.25%.

Authors

Person: Not available Organization: arXiv
A Knowledge Graph-Based Hallucination Benchmark for Evaluating ...

Sources

A Knowledge Graph-Based Hallucination Benchmark for Evaluating ... arxiv.org arXiv via serper

Referenced by nodes (1)

KGHaluBench concept