Friday, October 2, 2026
Sandboxing may not be sufficient to contain rogue agents, but the fact that AI labs have chosen not to implement good sandboxes (which we know how to do) gives me zero confidence that they will ever have success at preventing rogue agents (which no one knows how to do)