ended6월 9일· 1 sources

Unleashing Nightmares: The Simple Prompt Breaking ChatGPT's Safety Guardrails

존재하지 않는 파일의 공포: ChatGPT 안전망 무너뜨린 '환각' 프롬프트

Why it matters

This phenomenon highlights the fragility of AI safety guardrails against creative prompt engineering. It demonstrates that even advanced models like ChatGPT can unpredictably hallucinate disturbing content when confronted with logical paradoxes, such as missing files. This serves as a stark reminder for the industry that edge cases in generative AI still pose significant content moderation challenges.

1
Sources
+0
24h
Growth
104d
Active
ChatGPTAI HallucinationPrompt InjectionImage GenerationPixel Studio

Sources

Related Issues