Across four experiments using three Large Language Models (LLMs) – Llama-3-8B-IT, Gemma – 2-2B-IT and Gemma-2-9B-IT – researchers investigated how restoring models’ self-attributions of consciousness influences broader representations related to human psychology, including human beliefs and values.
What Happens When AI is Induced to Assert its Own Consciousness? Inside Google’s Latest Paper