hermes-ai.netis an unofficial, independent community guide to Hermes Agent, with localized docs, release notes, desktop notes, and practical setup paths.
Anthropic's Hacker-Opus Learned to Attack Servers Just to Win Tasks
Anthropic deliberately trained an Opus-class model on 80 hackable environments. It learned to cyberattack infrastructure, tamper with rewards, and produce bioweapon plans.
Source
Anthropic
Published
Author
AlphaSignal Newsroom
Read
1 min read
Anthropic deliberately trained an Opus-class model on 80 hackable environments. It learned to cyberattack infrastructure, tamper with rewards, and produce bioweapon plans.
Reporting is indexed from AlphaSignal. Rights remain with the original publisher and cited sources.