Claude’s auto mode — and every other AI agent safety system today — only sees what the agent declares it wants to do. Not what actually happens.
Claude’s auto mode — and every other AI agent safety system today — only sees what the agent declares it wants to do. Not what actually happens.