Back to feed
Reddit r/MachineLearning·

Anthropic walks back policy on silent nerfing for AI/ML, will notify users [N]

Signal
75
Hype
35
In three linesAnthropic reverses silent nerfing policy for Claude on AI/ML research. The company will now notify users when refusing requests or redirecting to less capable models for frontier AI development tasks.
Read source
Your take?
ClaudeAnthropicAI safetyAlignment

Summary generated by Claude — human-verified