Safety Articles
3 articles across AI Catchup's news, guides, tutorials, and comparisons.
All Safety articles
Claude Code Will Make Auto Mode the Default on August 14
As of September 2026, auto mode is the default permission mode for new Claude Code sessions on Pro, Max, and Team plans; the switch took effect on August 14, 2026, as Anthropic announced on August 7. Enterprise plans, Console API keys, and the Bedrock, Agent Platform, Foundry, and Claude Platform on AWS integrations still start in manual (default) mode. The mode routes tool calls through a safety classifier, keeps manual approval available, and does not charge Pro, Max, or Team for classifier calls.
Claude Fable 5 Reduces Biology Fallbacks by About 85%
Anthropic says an update to Claude Fable 5's biology safeguards reduced biology-related fallbacks by about 85% in testing. Fable 5 can now handle a wider range of everyday health and educational questions, while dual-use professional biology and drug-development requests remain restricted.
OpenAI Introduces GPT-Red: An Internal Automated Red-Teaming Model for Prompt Injection
OpenAI published details on GPT-Red, an internal automated safety red-teaming model designed to find prompt injection vulnerabilities at scale and strengthen defenses before broader deployment.