The headline today: Claude Code is switching to auto mode by default for Pro, Max, and Team plans starting August 14. Anthropic’s evals claim 89% harmful-action blocking (vs 13.6% for human reviewers) and 0/720 prompt-injection attacks succeeded against Fable 5 / Opus 5 / Sonnet 5 in a third-party eval — though Simon Willison’s analysis keeps a healthy dose of skepticism about the attack surface, including malicious packages in test suites.

Agent frameworks & tooling

  • Auto mode is now the default in Claude Code for Pro, Max, and Team plans — Starting Aug 14, Claude Code sessions will default to auto mode. Anthropic’s evals claim 89% harmful-action blocking vs 13.6% for human reviewers, and 0/720 prompt-injection attacks succeeded against Fable 5 / Opus 5 / Sonnet 5 in a third-party eval. Simon Willison’s analysis includes both the data and healthy skepticism (the “malicious package in test suite” attack surface remains). (Techmeme · simonwillison.net)

Industry

  • Software Giant SAP Stops Most Travel and Hiring Because of AI’s Soaring Cost — Internal SAP email obtained by 404 Media shows the freeze is still in effect since July, with exceptions only for AI-related hires/travel. An employee notes the company is rolling out a new internal AI tool “which massively increases costs.” Real data point on the tokenpocalypse hitting enterprise budgets. (HN 32pts · 404 Media)
All gathered items - what was cut and why (8)