Search papers, labs, and topics across Lattice.
This paper introduces a multi-agent reasoning framework for stance detection that enhances the robustness of predictions by shifting from label-level aggregation to reasoning-level synthesis. By employing a Manager-Worker architecture, the system adaptively allocates worker agents based on input complexity, allowing for diverse perspectives and reasoning explanations without directly emitting stance labels. The framework demonstrates significant performance improvements, particularly on implicit and context-dependent stances, achieving a Macro-F1 score of 86.07 on COVID-19 and 82.90 on SemEval-2016 datasets, highlighting the effectiveness of adaptive reasoning in complex stance detection scenarios.
Shifting from label aggregation to reasoning synthesis, this framework reveals that nuanced stance detection thrives on diverse agent perspectives rather than simplistic voting.
Stance detection requires identifying an author's position toward a target, often from short-form texts where stance is implicit, indirect, or rhetorically framed. Although large language models (LLMs) achieve strong performance on this task, single-pass prompting can be brittle when multiple interpretations are plausible. Existing aggregation strategies, such as majority voting or self-consistency, improve robustness by combining labels, but they discard the intermediate reasoning needed to resolve conflicting interpretations. We introduce a multi-agent reasoning framework with adaptive worker allocation for stance detection that shifts aggregation from label-level voting to reasoning-level synthesis. The framework employs a Manager-Worker architecture in which a Manager adaptively allocates a variable number of Worker agents based on input complexity. Each Worker analyzes the input from a distinct perspective and produces a reasoning-only explanation without emitting a stance label; the Manager then synthesizes these explanations to produce the final prediction. We evaluate the proposed framework on SemEval-2016, P-Stance, and COVID-19 Stance using Llama, Mistral, and Gemini. Results show that the framework yields the largest gains on implicit and context-dependent stance cases, achieving 86.07 Macro-F1 on COVID-19 and 82.90 on SemEval-2016, while remaining competitive on more explicit stance datasets such as P-Stance. These findings suggest that adaptive reasoning-level aggregation is most beneficial when stance cannot be reliably inferred from surface cues alone.