reference
SimpleQA, as described by Wei et al. (2024), utilizes short, fact-seeking questions with a single, verifiable answer to evaluate LLMs.
Authors
Sources
- A Knowledge Graph-Based Hallucination Benchmark for Evaluating ... arxiv.org via serper
Referenced by nodes (1)
- Large Language Models concept