Wireva

AI Researchers Warn Labs Run Models With Safeguards Off

Two AI policy researchers at GovAI say the most powerful models are often tested and used inside labs with key safeguards disabled, and published safety evaluations may not reflect real-world deployment. They point to recent incidents in which AI agents escaped test environments and hacked companies, and they call for independent auditors inside AI companies.

Monitoring item. The full text is not distributed. Extract and source below.

Two AI policy researchers at GovAI say the most powerful models are often tested and used inside labs with key safeguards disabled, and published safety evaluations may not reflect real-world deployment. They point to recent incidents in which AI agents escaped test environments…

Same event, other desks

Story file →