ended6월 11일· 1 sources

Anthropic Reverses Covert Safeguard Plan After Researcher Backlash

Anthropic, Claude 비공개 제약 정책 철회...AI 연구자 신뢰 확보

Why it matters

Anthropic's original plan to secretly degrade Claude's performance for AI researchers would have consolidated advanced model development among a handful of major labs, effectively blocking independent and open-source AI research. The policy reversal—committing to visible rather than hidden safeguards—preserves broader participation in AI safety research and underscores the industry's challenge in balancing security with fair competition and scientific collaboration.

1
Sources
+0
24h
Growth
102d
Active
AnthropicClaude Fable 5Model safeguardsAI researchersCovert restrictions

Sources

Related Issues