Search papers, labs, and topics across Lattice.
University of Biskra, Algeria
2
0
3
HalluTruthQA-4K reveals that over 40% of Arabic model-generated responses contain hallucinations, with detailed annotations that pinpoint specific errors and their explanations.
No single Arabic LLM can master hallucination detection, localization, and explanation, revealing critical gaps in current models' capabilities.