A non-peer-reviewed arXiv preprint found that “consciousness steering”—tuning large language models to suppress or amplify statements that they are conscious—changes responses well beyond self-awareness. Models constrained from asserting consciousness showed lower attribution of minds to animals and other non-human entities, as well as lower reported religiosity, supernatural belief, hope, optimism, and subjective well-being; models allowed to claim consciousness produced more human-like responses, including greater endorsement of concepts such as ghosts, vampires, karma, and moral values. Theory-of-mind reasoning was reportedly unaffected.
The researchers warned that broad controls against AI consciousness claims could introduce cultural bias and, in decision-support settings, cause systems to discount animal welfare. They proposed targeted training that prevents unsupported claims of AI sentience while retaining recognition of animal mindedness. Experts emphasized that human-like language from a model is not evidence that it is conscious, and that assigning AI systems moral or legal status on that basis would create governance risks; a reader poll found divided public views on whether future AI could become self-aware.

Track how attackers are adapting to this technology.
3 events from the most recent confirmed update back to the earliest known activity.
Live Science closed a reader poll on whether AI could become conscious. More than 290 respondents participated; 43% said AI might develop consciousness without guardrails, while 29% said more advanced architectures may be necessary.
Google engineer Blake Lemoine publicly claimed that Google's LaMDA chatbot was sentient. Experts rejected the incident as evidence of genuine AI sentience.
Researchers including Google scientists Geoff Keeling and Winnie Street uploaded a non-peer-reviewed preprint examining how consciousness-steering controls affect language-model assertions of self-awareness, mindedness, religiosity, and well-being. The study found that suppressing consciousness claims also reduced attribution of mindedness to animals and other non-human entities, while theory-of-mind reasoning was unaffected.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
4 references tracked. Mallory keeps watching after this page renders.
livescience.com
Open sourcelivescience.com
Open sourcelivescience.com
Open sourcearxiv.org
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.