Search papers, labs, and topics across Lattice.
This paper explores the critical intersection of hate speech detection (HSD) and privacy, highlighting that current HSD systems may inadvertently compromise user privacy by encoding authorship. The authors introduce the concept of a privacy-HSD trade-off, emphasizing the necessity of balancing effective hate speech detection with the protection of individual privacy rights. Through benchmarking various text privatization methods alongside their novel AgnoSpeech technique, they demonstrate that achieving this balance is challenging yet possible, underscoring the need for further research in this area.
Balancing hate speech detection and user privacy is not just a challenge鈥攊t's a necessary trade-off that could redefine online safety standards.
Hate speech is a real and timely threat that affects a large portion of online users, especially youth and minority groups. While building reliable and robust automatic hate speech detection (HSD) systems is paramount, we argue that this must also be balanced with the individual right to privacy. Exploring the intersection of HSD and privacy, we demonstrate that HSD systems might unintentionally achieve performance at the cost of encoding authorship, posing a threat to privacy. Building on these findings, we establish the notion of a privacy-HSD trade-off, which demands a careful balance. We benchmark a series of text privatization methods, as well as our newly proposed domain-specific AgnoSpeech technique, showing that balancing privacy and HSD is difficult but feasible. The findings make a strong case for more research on the trade-offs between privacy and HSD, both of which have tangible implications for the safeguarding of online participation.