OpenAI instructs ChatGPT models to cease mentioning 'goblins' after finding 66.7% of mentions linked to a specific personality.

OpenAI has mandated changes to ChatGPT models following a discovery that a specific personality was responsible for the majority of 'goblin' mentions. This action highlights concerns over the inappropriate use of metaphors by the AI.

Sources:
BBCYourStory.com+1
Trending 4h ago
Tab background
Sources: YourStory.comNDTV Profit
OpenAI has taken measures to prevent its AI tools from referencing 'goblins' after discovering that 66.7% of the mentions were tied to a specific personality trait.
The rise in goblin references was noted after the launch of GPT-5.1 in November, with a startling 175% increase in mentions compared to previous versions. Additionally, mentions of 'gremlin' also rose by 52%.
Despite this personality being used in only 2.5% of responses, it disproportionately affected the language patterns, accounting for 76.2% of preferences within that setting.
In response to these findings, OpenAI acted swiftly to rectify the issue, retiring this 'Nerdy' personality after it became clear that the AI's reward system had unintentionally favored creature-related language. The goblin phenomenon was further reinforced during testing of the new GPT-5.5 model, prompting additional scrutiny.
OpenAI confirmed that they had removed the 'goblin-affine reward signal' from their training data to ensure that references to goblins would no longer appear inappropriately.
This removes the unintended propensity for peculiar terms to emerge in a computer-generated conversation, thus maintaining a level of control over AI characteristics.
Sources: BBC
OpenAI has instructed its ChatGPT models to stop referencing 'goblins' after research revealed that 66.7% of such mentions were connected to a specific personality. This quirk, noted after the GPT-5.1 launch, saw a 175% increase in goblin mentions and prompted swift corrective action.
Section 1 background
The Headline

OpenAI restricts goblin mentions in ChatGPT

Key Facts
  • OpenAI had to instruct its AI tools to stop talking about goblins due to random mentions in responses.BBC
  • OpenAI discovered that a particular personality style was responsible for 66.7% of all 'goblin' mentions.BBC
  • Although only 2.5% of responses used the 'Nerdy' personality, it accounted for 66.7% of all 'goblin' mentions.YourStory.com
  • Responses containing words like 'goblin' or 'gremlin' were preferred in 76.2% of cases within that personality setting.YourStory.com
Section 2 background
Other Updates

Further developments from OpenAI

Key Facts
  • OpenAI added instructions to curb Codex's 'strange affinity for goblins' in its blog post.BBC
  • This would lead to those rewarded behaviours patterns being increasingly reproduced by AI.NDTV Profit1
Section 3 background
Background Context

Background on goblin mentions issue

Key Facts
  • The reason for this digital goblin incursion stems from an incentive system set up as a part of its personality customisation feature.NDTV Profit1
  • During testing of GPT-5.5 in Codex, engineers noticed the pattern again.YourStory.com
  • The creatures were being used as metaphors to a notably common degree, even in situations that had no connection to them.NDTV Profit1
Article not found
CuriousCats.ai

Article

Source Citations