Search papers, labs, and topics across Lattice.
This paper introduces UFPR-PEs, a benchmark dataset designed to evaluate face recognition bias using public videos of Brazilian politicians, annotated with self-declared race/color labels based on the Brazilian census taxonomy. The study reveals significant variations in recognition performance based on image quality and emphasizes the importance of analyzing subgroup gaps in conjunction with visual difficulty. By providing a reproducible framework for assessing demographic reliability in face recognition systems, UFPR-PEs addresses critical challenges in ensuring fairness and robustness in AI applications under real-world conditions.
Recognition performance in face recognition systems can vary dramatically based on image quality, revealing hidden biases that challenge existing benchmarks.
While face recognition systems are widely deployed, ensuring their demographic reliability and robustness under uncontrolled visual conditions remains a critical challenge. To bridge this gap, we present UFPR-PEs, a benchmark for face recognition bias evaluation using public videos of elected Brazilian politicians annotated with official self-declared race/color categories. The dataset adopts the Brazilian census taxonomy, including the parda category, which has no direct equivalent in the U.S.- or Europe-centric schemas commonly used in prior benchmarks. Our benchmark is built from compressed public video and preserves difficult samples so that performance can be analyzed under realistic conditions. We describe the construction pipeline, report dataset statistics, and evaluate face recognition performance across verification and (closed- and open-set) identification settings, including subgroup analysis by race/color and difficulty level. The results show that recognition performance varies substantially with image quality, and that subgroup gaps must be interpreted jointly with visual difficulty rather than in isolation. Overall, UFPR-PEs provides a reproducible and demographically grounded setting for studying face recognition bias under challenging public video conditions.