Search papers, labs, and topics across Lattice.
This paper introduces GPE, a novel multi-domain benchmark and evaluation framework designed to assess the robustness of fact verification methods against controllable GEO-style poisoning attacks. The study reveals that existing benchmarks fail to capture the vulnerabilities introduced by manipulated evidence, highlighting significant degradation in performance and efficiency trade-offs when subjected to adversarial conditions. By conducting experiments across various verification methods, the authors confirm that GPE provides critical insights into the limitations of current evaluation practices in the face of manipulated information sources.
Fact verification methods can degrade significantly when faced with controlled evidence poisoning, revealing vulnerabilities that traditional benchmarks overlook.
Large language models increasingly use search tools to retrieve up-to-date information, introducing a new attack surface in which retrieved documents can be manipulated. This risk is amplified by the development of generative engine optimization, which can make selected content more likely to be retrieved, cited, and adopted by models. Existing fact-verification benchmarks and evaluation frameworks do not provide the controlled evidence environments needed to assess robustness against GEO poisoning. We therefore propose GPE, which consists of a multi-domain fact-verification benchmark and an evaluation framework for controlling evidence sources and poisoning ratios. Experiments across multiple verification methods and poisoning attacks demonstrate that GPE exposes robustness degradation and efficiency trade-offs that cannot be observed through clean evaluation alone, confirming the need to evaluate fact verification under adversarial evidence environments.