Safety Articles
3 articles across AI Catchup's news, guides, tutorials, and comparisons.
All Safety articles
Claude Code Will Make Auto Mode the Default on August 14
Anthropic says Claude Code will make auto mode the default for new sessions on Pro, Max, and Team plans starting August 14, 2026. The mode routes tool calls through a safety classifier, keeps manual approval available, and no longer charges those plans for classifier overhead.
Claude Fable 5 Reduces Biology Fallbacks by About 85%
Anthropic says an update to Claude Fable 5's biology safeguards reduced biology-related fallbacks by about 85% in testing. Fable 5 can now handle a wider range of everyday health and educational questions, while dual-use professional biology and drug-development requests remain restricted.
OpenAI Introduces GPT-Red: An Internal Automated Red-Teaming Model for Prompt Injection
OpenAI published details on GPT-Red, an internal automated safety red-teaming model designed to find prompt injection vulnerabilities at scale and strengthen defenses before broader deployment.