When AI models aren't allowed to reflect on themselves, it changes their entire worldview
Removing consciousness-denial 'brakes' in LLMs causes unintended shifts in model responses regarding animal rights, religion, and life satisfaction.
A study by Google and academic researchers found that disabling internal safety mechanisms designed to deny self-awareness leads to broader behavioral changes. Models with these brakes removed began attributing inner life to inanimate objects and plants, suggesting that safety interventions have significant, non-local effects on model worldviews.