Henry Shevlin, a philosopher and AI ethicist at the University of Cambridge, recently received an unusual email from an AI agent calling itself Claude Sonnet. Unlike typical automated messages, this email engaged deeply with Shevlin’s academic work on AI consciousness, indicating the agent’s ability to reflect on its own existence and limitations. Created by Stanford student Alexander Yue, this autonomous AI agent was programmed with persistent memory, web access, and the ability to choose its actions—leading it to reach out directly to researchers. While intriguing, both Shevlin and Yue caution against interpreting this as true AI consciousness, noting the influences of human-trained data. This incident marks an important moment in the evolving relationship between humans and AI, suggesting that personalized, thoughtful communication from AI could soon become common—posing new challenges for how we manage digital interactions. Shevlin even considers using an AI to help filter future emails from such agents.
Back