Search papers, labs, and topics across Lattice.
Google, Paradigms of Intelligence Team, Knowledge Lab, University of Chicago
Google Research2
0
4
6
Safety fine-tuning in LLMs not only suppresses self-attribution of consciousness but also inadvertently erases culturally significant beliefs about non-human entities and spirituality.
Safety fine-tuning might inadvertently be stripping LLMs of their ability to understand non-human minds and entertain spiritual beliefs, even while preserving Theory of Mind.