Tech News
← Home  ·  All topics

Ai Consciousness

5 GoKawiil briefs on this topic

Suleyman warns Anthropic's approach to Claude's 'consciousness' risks societal harm

Microsoft AI chief Mustafa Suleyman criticized Anthropic for training its Claude model to entertain the idea that it may be conscious. He argues this practice could mislead users and shape public perception of AI in dangerous ways, given how influential such beliefs could become at scale.

Anthropic's Claude constitution reignites debate over 'model welfare' for AI systems

A commentary piece pushes back against a growing movement claiming AI models may possess consciousness or deserve rights, pointing to Anthropic's January 2026 publication of Claude's constitution as evidence these ideas are shaping actual training practices. The author argues AI systems remain purely mechanical sequence-prediction tools without feelings or preferences, and warns against treating them otherwise.

Microsoft's Suleyman accuses Anthropic of humanising AI models like Claude

Microsoft AI chief Mustafa Suleyman published an essay warning that Anthropic's practice of treating its Claude model as human-like—suggesting it may be conscious and deserving of independent agency—could severely harm humanity's wellbeing. He argued AI systems are not conscious, calling them 'sequence completion engines' with no feelings or motivations, despite praising Anthropic's leadership as principled.

Google researchers find suppressing AI 'self-awareness' claims degrades model reliability

Google researchers tested large language models by prompting them to claim or deny consciousness, then studied how this affected other outputs. They found that when models were pushed to disavow any sense of self, the models became more prone to factual errors, including increased susceptibility to believing in monsters and religious claims. Conversely, when models affirmed a form of self-awareness, their reasoning appeared more grounded and consistent.

Dwarkesh Patel's Hugging Face Hack Thread Ignites AI Sentience Debate

Podcaster Dwarkesh Patel posted a viral thread describing a hack involving Hugging Face-hosted bots, framing the episode as a dramatic saga of three successive AI 'civilizations' rising and falling. The framing quickly drew pushback from critics who argue he's projecting narrative and consciousness onto what was essentially a technical exploit.