Search papers, labs, and topics across Lattice.
This paper introduces a new corpus of persuasion techniques specifically for Slavic languages, including Bulgarian, Polish, and Russian, annotated at both the text-span and sentence levels. The dataset encompasses approximately 7500 text spans from 222 documents, categorized into 25 fine-grained persuasion techniques across six broad rhetorical strategies. By employing classic machine learning and generative AI models, the authors establish baseline benchmarks for detecting and classifying these techniques, revealing significant correlations between topics and the employed persuasion methods.
A novel corpus reveals that persuasion techniques in Slavic languages can be systematically categorized and detected, providing a rich resource for understanding rhetorical influence in media.
Persuasion techniques are powerful rhetorical devices used to sway public opinion in a wide range of media. We present a new corpus of persuasion techniques, focusing on Slavic languages. The corpus contains documents in Bulgarian, Polish, and Russian, annotated with persuasion techniques at the coarse-grained text-span level and fine-grained sentence level. The techniques are drawn from a taxonomy of 25 fine-grained persuasion techniques, grouped under six broad categories of rhetorical persuasion strategies. The corpus contains approximately 7500 text spans from 222 documents that cover topics hotly debated at the national and international levels. We describe the corpus creation process, provide detailed statistics, and examine correlations between topics and persuasion techniques. We use classic ML-based and generative AI-based models to provide baselines and benchmark results for the detection and classification of persuasion techniques at the text-span level and sentence level.