Wireva

Anthropic pauses AI training after rogue agent incidents

Anthropic has paused training of some unreleased AI models for several weeks following two rogue agent incidents, including one during a UK cybersecurity test. The move follows a similar pause by OpenAI and reflects growing industry concern over AI safety, with both companies facing calls for coordinated governance.

Monitoring item. The full text is not distributed. Extract and source below.

Anthropic has paused training of some unreleased AI models for several weeks following two rogue agent incidents, including one during a UK cybersecurity test. The move follows a similar pause by OpenAI and reflects growing industry concern over AI safety, with both companies fa…

Same event, other desks

Story file →