Search papers, labs, and topics across Lattice.
This paper introduces the concept of affective safety in AI, highlighting the overlooked risks associated with AI systems' interactions with human emotions. By developing a comprehensive taxonomy of affective harms鈥攕uch as affective self-alienation, fairness and bias harms, and relational harms鈥攖he authors reveal how these issues recur across various AI systems, indicating a need for a unified approach. The study emphasizes the inadequacy of current safety frameworks in addressing these concerns and calls for dedicated strategies to mitigate the cumulative and relational impacts of AI on human emotional well-being.
Affective safety is a critical yet neglected dimension of AI risk that could reshape our understanding of AI's impact on human emotional health.
AI safety research has focused predominantly on epistemic and physical harms (e.g., misinformation, bias, system reliability) while the risks that arise from AI systems' engagement with human emotional life have remained fragmented and undertheorised. We propose affective safety as a unified class of AI safety concerns grounded in the fact that humans are affective beings. We develop a taxonomy of affective harms and identify recurring harm types: (1) affective self-alienation, (2) fairness and bias harms, and (3) relational harms. We show that their recurrence across system types reflects structural properties of how AI systems engage with human emotion and survey the current safety landscape and show that existing frameworks address affective safety either narrowly or not at all. We conclude by identifying the technical and regulatory challenges specific to this class of harms and argue that affective safety requires dedicated frameworks that engage with cumulative, relational, and identity-level effects.