- OpenAI works to stop ChatGPT from generating 'sex crime scene' images, as researchers highlight the AI's ability to produce harmful content.
- OpenAI stated, "After investigating this trend, we've introduced additional safeguards against this type of prompt."
- Researchers noted that the AI produced a range of gory and sexualised images of "its own volition".
- One generated image was titled "Grim crime scene aftermath", depicting disturbing content.
- OpenAI has implemented automated systems and human review to identify and block harmful material.
- OpenAI's policies prohibit sexual violence, non-consensual intimate content, and attempts to bypass its safeguards.
OpenAI is actively addressing concerns regarding ChatGPT's ability to generate harmful content, including sexualized images and graphic violence. Researchers revealed that the AI could produce such material with simple prompts, prompting OpenAI to implement new safeguards.12
The latest public version of ChatGPT was reported to generate sexualized images and graphic violence, with one prompt resulting in a depiction titled "Grim crime scene aftermath".5
OpenAI stated, "After investigating this trend, we've introduced additional safeguards against this type of prompt," emphasizing its commitment to preventing the generation of inappropriate content. The company's policies explicitly prohibit sexual violence, non-consensual intimate content, and child sexual abuse material, as well as attempts to bypass its safeguards.37
A spokesperson for the Department for Science, Innovation and Technology acknowledged that while "safeguards in AI models are improving, there is more to do", highlighting the ongoing challenges in ensuring AI safety and ethical use.
As AI technology continues to evolve, the need for robust safeguards becomes increasingly critical to prevent the dissemination of harmful content.
“OpenAI is actively addressing concerns over ChatGPT's capability to generate graphic and sexualised images. The company has implemented safeguards but acknowledges that challenges remain in preventing harmful content.”
