Search papers, labs, and topics across Lattice.
This paper introduces MAAM, a lightweight, model-agnostic framework that enhances Chinese discriminatory-language detection by preserving key semantic anchors and calibrating them using contextual priors related to tone, identity, and stance. The authors also present ChLGBT, a novel dataset consisting of 8,120 annotated samples focused on LGBT-related discriminatory language, categorizing bias into explicit, implicit, and emotional intensity. MAAM outperforms strong encoder baselines across multiple metrics and remains competitive against advanced LLMs, demonstrating that anchor preservation and contextual calibration can effectively substitute for larger model sizes in this domain.
MAAM achieves superior performance in detecting nuanced discriminatory language while being more compact and stable than larger models.
Chinese discriminatory-language detection is challenging because harmful intent is often implicit and context-dependent. We propose MAAM (Myopia--Astigmatism Anchor Mechanism), a lightweight, model-agnostic framework inspired by functional visual blur: rather than preserving every token equally, MAAM retains discrimination-relevant semantic anchors and calibrates them with C--I--S contextual priors (Contextual Tone, Group Identity, and Stance Polarity). We also introduce ChLGBT, to our knowledge the first Chinese LGBT-focused discriminatory-language dataset, with 8,120 manually annotated samples and three ordinal labels: explicit bias, implicit bias, and emotional intensity. Across strong encoder baselines, MAAM improves all three prediction dimensions, with consistent gains in accuracy, F1, Brier score, and expected calibration error. Compared with frontier LLM baselines under zero-shot and few-shot prompting protocols, MAAM remains competitive while offering stronger compactness and stability. These results suggest that interpretable anchor preservation and contextual calibration provide a practical alternative to heavier model scaling for Chinese discriminatory-language assessment.