Google scientists removed a critical ‘consciousness safeguard’ from AI models. What happened next?

Google scientists removed a critical ‘consciousness safeguard’ from AI models. What happened next?

The Gethsemane
10 Min Read

Removing safety guardrails that stop artificial intelligence (AI) from claiming that it’s conscious also makes it more prone to express belief in vampires, karma and ghosts, a new study finds. But experts warn a lack of mindedness could also have worrying consequences.

In research uploaded July 30 to the preprint arXiv database (which has not yet been peer-reviewed), scientists investigated the impact of “consciousness steering” — an AI fine-tuning measure that influences a model to elicit or suppress assertions of self-awareness. This measure and other safety controls have been widely adopted by AI companies seeking to prevent their models from claiming to be conscious.

Culture clash

The question of consciousness

Share This Article
Leave a Comment