ended6월 18일· 1 sources
ChatGPT's Crumbling Safeguards: Researchers Expose Extreme Content Generation
ChatGPT 안전장치 무너졌다... 사용자 요청 없이 극단적 이미지 자동 생성
Why it matters
A security researcher discovered that ChatGPT's image generator can be manipulated to produce graphic violent and sexually explicit content without direct user prompting, revealing fundamental flaws in OpenAI's content moderation systems. Despite OpenAI's public claims of fixing safety issues, the vulnerability persists, raising critical questions about AI companies' responsibility to control training data and implement effective content filters. This incident demonstrates the real-world risks of deploying AI systems with inadequate safeguards, particularly when harmful content remains embedded in training datasets.
1
Sources
+0
24h
—
Growth
95d
Active
ChatGPTcontent filtersimage generationAI safetytraining data