Search papers, labs, and topics across Lattice.
To disentangle an LLM's default cultural priors from its capacity for contextual adaptation, this work evaluates six instruction-tuned models on a 12-culture, 304-item forced-choice benchmark using a four-level context steering gradient. Existing cultural evaluations rely on single-ground-truth accuracy, failing to characterize real-world scenarios where multiple culturally grounded choices are simultaneously valid. Across all evaluated models, default cultural priors skew heavily toward the US and UK (~35% of selections across 12 cultures), while prompt-based steering counterintuitively widens the representation gap between high- and low-resource cultures.
Prompt-based personalization cannot fix cultural bias: steering LLMs with cultural context actually widens the disparity between dominant and underrepresented cultures, even when injecting explicit cultural facts.
Large language models (LLMs) are increasingly deployed in globally used assistants, yet their default choices in culturally grounded everyday situations can systematically favour some cultures over others, affecting localisation, user trust, and equitable behaviour. Existing cultural benchmarks evaluate accuracy against a single "correct" answer, making it difficult to characterise an LLM's cultural preference prior when multiple culturally grounded responses are all valid; they also conflate default preferences with context-driven adaptation. We propose DiSCo, a distribution-first forced-choice evaluation framework that isolates default cultural priors and tests steerability via a four-level context gradient (C0--C3). Using DiSCo-Bench (304 items) derived from BLEnD spanning 12 cultures, we evaluate six diverse instruction-tuned LLMs. Default priors are heavily concentrated, with UK and US together absorbing approximately 35\% of all selections despite representing only 2 of 12 cultures. Most critically, prompt-based steering consistently widens the selection gap between high- and low-resource cultures, and injecting explicit cultural facts produces negligible distributional disruption, confirming that cultural preference bias cannot be resolved through prompt-based personalisation alone.