Skip to main content
AI Safety & Governance

Anthropic Optimizes Claude Fable 5 Model’s Biosafety Mechanisms, Reducing False Blocks by 85%

Anthropic updated Claude Fable 5’s biosafety safeguards yesterday. By adjusting its safety classifiers, the model can more accurately distinguish harmless questions from high-risk ones. Official tests show that refusals for biology-related questions have decreased by 85%, allowing users to handle a broader range of biological tasks while preventing malicious use. #AI Safety#

Anthropic Optimizes Claude Fable 5 Model’s Biosafety Mechanisms, Reducing False Blocks by 85%

Anthropic, a U.S. artificial intelligence startup, said in a blog post yesterday that it is updating the Claude Fable 5 model’s biosafety safeguards, significantly reducing false blocks. When Fable 5 users ask biology-related questions, the system is now less likely to downgrade its response.

Anthropic Optimizes Claude Fable 5 Model’s Biosafety Mechanisms, Reducing False Blocks by 85%

According to Anthropic’s official test data, this update reduced refusals for biology-related questions by 85%, and Fable 5 can now handle a broader range of biological tasks.

Anthropic said that some malicious users use the Fable 5 model for high-risk activities such as biological weapons research. To prevent this, the company specifically configured safety classifiers for Claude to detect whether users are asking Fable 5 to perform protected biological tasks or generate harmful content.

Anthropic Optimizes Claude Fable 5 Model’s Biosafety Mechanisms, Reducing False Blocks by 85%

Anthropic has now adjusted how its safety classifiers work so that they can more accurately distinguish harmless questions from high-risk ones.

However, to guard against ongoing AI risks, Fable 5 will still restrict requests related to specialized biological research and drug development.