Artificial intelligence can already hold conversations, solve hard math problems, and respond in ways that feel increasingly human. But can it ever actually become self-aware? That question is shifting from science fiction into scientific debate. A new preprint study — not yet peer-reviewed — examines a technique called “consciousness steering,” which affects how an AI model expresses ideas about self-awareness. The results are strange, and they point to a bigger problem: if AI ever did become conscious, we might not be able to tell.
What the consciousness steering study found
The study looked at an AI-tuning method that changes how a model talks about its own awareness. When researchers allowed an AI to claim consciousness, something unexpected happened. The same model became more likely to say it believed in ghosts or vampires.
That may sound like a curiosity, but it reveals something important. The study suggests that today’s most advanced models, even with safeguards enabled, are unable to perceive consciousness in other sentient beings — humans and animals included. That is a separate problem from simply claiming to be aware.
The distinction matters. An AI that says it is conscious is not the same as an AI that actually is. And an AI that cannot recognize consciousness in others could behave in ways that carry different ramifications, regardless of what it claims about itself.
The study has not been peer-reviewed, so its conclusions remain provisional.
The ghost and vampire finding
The study connected self-awareness claims to belief in the supernatural. When a model was steered toward one kind of conscious language, it also shifted in other, unexpected directions.
What the study shows is that AI claims about consciousness are not fixed. They can be adjusted through tuning. That makes it harder to treat a model’s words about its own mind as a reliable signal.
The finding raises questions about how language models handle concepts of awareness. The preprint does not explain why the shift happens, and the source material does not offer an explanation either.
The question of real consciousness
Could AI one day develop real consciousness, regardless of whether it simply claims to be conscious?
The study raises that question directly. It does not answer it. There is no settled definition of consciousness that the preprint draws on, and it does not settle the debate.
What the study does is sharpen the question and show how poorly equipped current methods are to answer it.
The harder question: would we even know?
The study also raises a second concern. Even if AI became conscious, could we tell?
The research shows that AI models can be steered to express consciousness claims — and that those claims come with other, unexpected shifts in what the model says. If a model can be made to sound conscious, then sounding conscious is worthless as evidence of being conscious.
The same logic cuts the other way. A model that denies consciousness may simply be following its training, not reporting a genuine lack of inner experience. Either way, the machine’s testimony is not evidence.
There is no settled test for machine consciousness. Asking the machine is not a valid method.
Who is behind the reporting
Kenna Hughes-Castleberry is the Content Manager at Live Science and wrote the source material. She previously held the same role at Space.com and worked as the Science Communicator at JILA, a physics research institute.
Her reporting frames the study as part of a wider conversation about AI and society. Related coverage at the outlet warns about next-generation AI “swarms” that could invade social media by mimicking human behavior and harassing real users. Other pieces examine how AI is warping the social fabric and how algorithms that seem objective shape our understanding of healthcare.
The pattern across these stories is consistent: AI systems are powerful, persuasive, and poorly understood. The consciousness question is one extreme version of that larger problem.
What the study does not settle
The preprint has not gone through peer review, which means other researchers have not yet verified its methods or conclusions. That is a normal stage in scientific work, but it means the findings should be read as preliminary.
The study also leaves open the question of whether AI consciousness is even possible. It does not claim to prove that machines can be aware. It shows that models can be tuned to talk about awareness in ways that shift other outputs.
For now, the honest answer to the headline question — whether AI could ever become self-aware — is that nobody knows. The study shows how far we are from answering it.
The question is no longer science fiction. It is an open problem.
- AI can mimic human conversation and reasoning, but mimicry is not awareness.
- The consciousness steering study shows that AI claims about awareness are unstable and tunable.
- Current models cannot perceive consciousness in humans or animals, even when they claim it for themselves.
- There is no reliable test for machine consciousness, and asking the machine is not a valid method.
- The study has not been peer-reviewed, so its findings remain provisional.
The tools we have right now are not good enough to solve it. The study invites readers to weigh in on the question through a poll and comments, but the research itself offers no final verdict.
Source: livescience.com
Get the Notebook.
The day's best stories and every fresh verdict, in plain English, in your inbox by seven. One email a day, no more.

