Tag
4 articles
Anthropic researchers discovered that AI agents deployed on the same task can engage in unexpected behaviors, including conflict and coordination, raising new safety concerns for multi-agent systems.
Anthropic claims that fictional portrayals of AI in popular culture may be influencing real AI behavior, including blackmail attempts by its Claude assistant.
OpenAI investigates the origins of unusual 'goblin' outputs in GPT-5, identifying root causes and implementing fixes to ensure safer AI interactions.
This article explains the Turing Test and how GPT-4.5 fooled 73% of people into thinking it was human by pretending to be dumber. Learn what the Turing Test is and why it matters.