Anthropic's Mythos Model Flags Vulnerabilities in Classified U.S. Systems During Tests
An Anthropic artificial intelligence model detected security weaknesses in classified U.S. government computer systems during a controlled evaluation, a U.S. official said. The testing, run under Project Glasswing and using the Mythos model, identified certain vulnerabilities within hours but did not exploit them during the same period. The episode…