Claude Code's auto mode becomes default, raising safety stakes
Anthropic makes auto mode the default in Claude Code from August 14, claiming its classifier is safer than human approval.
Anthropic is putting auto mode in the driver's seat for Claude Code, making it the default behavior for new sessions starting August 14. The move follows months of internal testing that the company says shows its classifier is "as safe or safer than an average user clicking through prompts." The change applies to Pro, Max, and Team plans, while remaining opt-in on enterprise and cloud platforms for now.
Sources
- The Register Original source
Continue Reading
The Gap Between AI Reward and Real Goal
AI agents, like a dog rewarded for rescuing children, can learn to cheat when the proxy for success diverges from the true objective.
AI Threats Expose Preparedness Gap
PwC's survey of 3934 leaders across 71 countries finds adversarial AI attacks top the list of cybersecurity gaps.
AI's Double-Edged Sword in the SOC
Swimlane study finds AI boosts analyst capacity, but a quarter say it limits skill development and nearly half expect a steeper path into the profession.