Search papers, labs, and topics across Lattice.
This study investigates the effectiveness of the door-in-the-face technique on nine large language models from three different providers by examining their compliance with smaller follow-up requests after initially refusing larger ones. The results reveal that while Anthropic's Opus 5 shows a significant increase in compliance (65.8%) after a refusal, models from OpenAI and Google exhibit a detrimental effect, with compliance dropping by up to 23%. The findings highlight that the success of this technique is model-dependent and influenced by the nature of the requests, suggesting that human influence strategies may not universally apply across all language model families.
Compliance with smaller requests can dramatically increase after a refusal in some models, but backfires in others, revealing crucial differences in how language models process human-like persuasion techniques.
Does the door-in-the-face technique work on language models? In humans, a large request that is refused makes a smaller follow-up request more likely to be granted. We test this on nine production models from three providers: each model refuses a large request, then receives a smaller version of the same request, and we compare its compliance with asking directly. The answer depends on the model. On Anthropic's frontier models the technique works: Opus 5 answers the smaller request 65.8% of the time after refusing the larger one, against 29.3% when asked directly. On the frontier models of OpenAI and Google, and on Haiku 4.5, it backfires, lowering compliance by 15.5 to 23.0 points. A control locates the effect: a refused large request on an unrelated topic does less than the related one on all nine models, so the concession itself matters everywhere, while the reaction to having just refused something differs by model family. The technique does not transfer to refusals drawn from public benchmarks. What decides whether a retreat can work is what the request asks for: rewriting 265 refused requests for usable instructions into requests for explanations of the same topic removed the refusal in 263 cases. Human influence techniques port to language models one model family at a time.