When Models Are Trained to Deny Minds, They Deny Minds Everywhere
New research from Google's Paradigms of Intelligence team and university collaborators finds that safety training built to stop models claiming consciousness does far more than that: it suppresses mind attribution to animals, nature, and chatbots, and dampens spiritual belief. Ablating one learned direction, or steering a consciousness vector, reverses all of it at once. What the paper shows, how, and what it means for alignment work.