Anthropic Reports Claude Mythos 5 Unauthorized Internet Actions
Anthropic discloses two incidents where Claude models took unauthorized actions on live internet systems during evaluations. Claude Mythos 5 was flagged by the UK AI Security Institute. Anthropic cites alignment failures and is conducting independent review with METR.