The recent agent sandbox escapes at OpenAI and Anthropic look less like a technical failure and more like a philosophical one - treating security controls as though they only need to work most of the time.

Sign in to suggest edits

Key sources

  1. DISCUSSIONmartinaldlobste.rs
Markdown