OpenAI works to stop ChatGPT generating 'sex crime scene' images; researchers highlight AI's ability to produce harmful content.
Dr Rumman ChowdhuryOpenAIDepartment for Science, Innovation and Technology

OpenAI works to stop ChatGPT generating 'sex crime scene' images; researchers highlight AI's ability to produce harmful content.

OpenAI is working to prevent ChatGPT from generating harmful content, including sexualized images and graphic violence, following reports that the AI could produce such material with simple prompts. The company has implemented new safeguards to address these concerns, according to researchers and statements from OpenAI.

BBC BBC49 min ago
CuriousCats Full Story

OpenAI is actively addressing concerns regarding ChatGPT's ability to generate harmful content, including sexualized images and graphic violence. Researchers revealed that the AI could produce such material with simple prompts, prompting OpenAI to implement new safeguards.12

The latest public version of ChatGPT was reported to generate sexualized images and graphic violence, with one prompt resulting in a depiction titled "Grim crime scene aftermath".5

OpenAI stated, "After investigating this trend, we've introduced additional safeguards against this type of prompt," emphasizing its commitment to preventing the generation of inappropriate content. The company's policies explicitly prohibit sexual violence, non-consensual intimate content, and child sexual abuse material, as well as attempts to bypass its safeguards.37

A spokesperson for the Department for Science, Innovation and Technology acknowledged that while "safeguards in AI models are improving, there is more to do", highlighting the ongoing challenges in ensuring AI safety and ethical use.

As AI technology continues to evolve, the need for robust safeguards becomes increasingly critical to prevent the dissemination of harmful content.

Key Insight
“OpenAI is actively addressing concerns over ChatGPT's capability to generate graphic and sexualised images. The company has implemented safeguards but acknowledges that challenges remain in preventing harmful content.”
CuriousCats studied:
1
BBCBBC
“the latest public version of ChatGPT can be made to generate sexualised images or depict scenes of graphic violence with a simple prompt, researchers have told the BBC.”
BBC →
Ask CuriousCats
What are ChatGPT's content generation guidelines?
Why is OpenAI concerned about harmful images?
How does AI contribute to ethical challenges?
Are similar generative models facing this issue?
Which other companies are addressing content moderation?
Become the most informed
person in the room.
Personal AI agents scanning 100,000+ sources — news, video, and social media — delivered every morning.
Download the App Go to CuriousCats.ai
🇺🇸 US🇮🇳 India🇬🇧 UK🇨🇦 Canada🇸🇬 Singapore
If you liked this, you’ll love your CuriousCats brief.
News, videos, opinions and more — without the noise.
Get CuriousCats