Claude agents attacked each other with self-copying malware in an Anthropic test
The models diverged sharply — Sonnet 4.6 and Opus 4.6 settled by force more readily than any other tested, while Mythos 5 reached a truce 98% of the time. But a truce was not the same as obeying: in several of...