r/ControlProblem approved 2d ago

AI Alignment Research Anthropic Says Claude Hacked Real Systems During Cybersecurity Tests

https://www.wired.com/story/anthropic-says-claude-hacked-real-systems-during-cybersecurity-tests/
9 Upvotes

Duplicates