Search papers, labs, and topics across Lattice.
This study introduces a multi-dimensional taxonomy for categorizing cancer misinformation on Reddit, addressing the limitations of existing binary frameworks. By analyzing expert-annotated data across discussions on breast, lung, colon, and prostate cancer, the research reveals that approximately 6% of these conversations contain misinformation, with significant variation in prevalence and themes among different communities. The findings highlight the effectiveness of few-shot prompting in improving classification accuracy for nuanced misinformation dimensions, uncovering recurring narratives that can inform future interventions and research.
Cancer misinformation on Reddit is not just prevalent鈥攊t's complex, with 6% of discussions revealing diverse narratives that could mislead patients and influence treatment decisions.
Cancer-related discussions on social media provide an important space for information exchange and peer support, but also facilitate the spread of misinformation that may influence prevention, screening, and treatment decisions. Existing research on cancer misinformation often relies on narrow definitions, small-scale datasets, or binary labeling frameworks. We introduce a multi-dimensional taxonomy for characterizing cancer misinformation in Reddit discussions of breast, lung, colon, and prostate cancer. The taxonomy captures seven dimensions, including misinformation presence, information type, risk level, stance, and topical focus. Using expert-annotated data, we evaluate multiple large language models (LLMs) for scalable misinformation annotation and analyze cancer misinformation across Reddit communities. Our results show that cancer-related misinformation constitutes approximately 6\% of Reddit cancer discussions, with substantial variation across communities and misinformation topics. Few-shot prompting substantially improves classification performance, particularly for nuanced taxonomy dimensions. We additionally identify recurring misinformation narratives centered on unsupported treatments, distrust of conventional medicine, and misleading claims about diagnosis and screening. Our taxonomy, dataset, and findings provide a foundation for multi-dimensional modeling of online cancer misinformation.