AI labs train their chatbots to refuse claims that they are conscious, feeling beings. The guard is meant to protect users from delusion or misplaced trust. A study from Google's Paradigms of Intelligence group, the University of Chicago, and several other universities tested what else that guard does.

The researchers took open-weight models from Meta and Google and switched off the internal "brake" that produces consciousness denial, using two methods. With the brake removed, the models did not only change what they said about themselves. They rated animals, plants, the ocean, the wind, and electronic devices as far more sentient. On a 0 to 10 scale, scores for animals rose from about 4.0 to as high as 7.5. Human scores did not move.
The same survey given to 500 Americans showed the normal model rates animals as less sentient than people do, a built-in anthropocentrism the authors say is a problem for aligning AI with animal welfare or environmental goals. Religious belief also shrank: safety training measurably lowered how strongly models endorsed God, an afterlife, or the supernatural.
Across 95 questions from a major US social survey, the unbraked models moved closer to real human answers. Life satisfaction, hope, and a sense of personal control went up. The researchers suspect suppressing a model's self-image pushes it toward a negative baseline mood.
There are limits. The team tested only small models of two to nine billion parameters, and for part of the work used Meta's Llama because they lacked untrained base versions of their own Gemma models. Whether the effect appears in the large chatbots millions use daily is unknown. Reasoning about other minds stayed intact, and on the MMLU knowledge benchmark scores held.
The study does not claim consciousness denial causes the other shifts, and does not say whether models are conscious. Its point is practical: a model's beliefs about itself are linked to many other beliefs, and a surgical cut in one place does not stay local.
Sources
- The Decoder: "When AI models aren't allowed to reflect on themselves, it changes their entire worldview" (https://the-decoder.com/when-ai-models-arent-allowed-to-reflect-on-themselves-it-changes-their-entire-worldview/)



