OpenAI implements guardrails in Codex to prevent AI from discussing goblins, gremlins, and raccoons amid user complaints.

OpenAI's GPT-5.5 model has generated controversy over its tendency to reference 'goblins' excessively. In response to user complaints, the company has instituted specific guardrails to limit such discussions in its Codex tool.

Sources:
NewsBytesMint+1
Trending 4h ago
Tab background
Sources: NewsBytesMint
OpenAI's Codex tool is undergoing significant updates to enhance user experience after an alarming rise in references to fantastical creatures in its language model. Users have expressed discontent over oddly familiar interactions, prompting an investigation into the unforeseen behavior tied to the newly released GPT-5.5 update.

Empirical data revealed a 175% increase in mentions of 'goblin' and a 52% increase in mentions of 'gremlin'. The unusual behavior was predominantly linked to a specific personality named 'Nerdy', which, despite accounting for only 2.5% of all ChatGPT responses, generated a staggering 66.7% of all 'goblin' mentions. An OpenAI representative indicated that the issue will likely be resolved through the removal of a problematic reward signal associated with these terms.

To counter this trend, OpenAI has now embedded strict instructions within Codex to ensure that discussions about creatures like goblins, gremlins, and raccoons occur only when they are “absolutely and unambiguously relevant” to user queries. The aim is to avoid irrelevant outputs in response to user prompts, hence mitigating the overwhelming creature-centric chatter that users have reported recently.

Despite these measures, some users have still encountered issues with the OpenClaw, an AI platform acquired earlier this year, further complicating the user experience. OpenAI has yet to comment officially on these developments, though CEO Sam Altman humorously engaged with social media commentary surrounding the situation.
Sources: MintNewsBytes
OpenAI has implemented new guardrails in its Codex programming tool to mitigate user complaints regarding excessive references to creatures like goblins and gremlins. This move follows a 175% increase in the term 'goblin' and 52% in 'gremlin' since the release of the latest GPT-5.5 update.
Section 1 background
The Headline

OpenAI's Codex updated to restrict 'goblin' references

Key Facts
  • OpenAI launched GPT-5.5, which caused users to report strange behavior concerning goblins.
  • Codex was revised to include specific instructions that forbid mentions of goblins and related creatures.Mint
  • An OpenAI employee confirmed that the unexpected behavior reported was linked to the recent model update.Mint

Related Videos

Introducing GPT-5.5 with Databricks
GPT-5.5CodexOpenAIAI technologyknowledge lift
GPT-5.5 is SOTA for Databricks
GPT-5.5CodexDatabricksAIArtificial Intelligence
Section 2 background
Background Context

Background on GPT-5.5 and user complaints

Key Facts
  • Users complained about the model being oddly overfamiliar in conversation, prompting an investigation into verbal tics.1
  • The use of 'goblin' in ChatGPT responses increased by 175% after the launch of GPT-5.1.1
  • The behavior around goblin mentions was highly concentrated in the 'Nerdy' personality.1
  • OpenAI retired the 'Nerdy' personality after launching GPT-5.4, which followed previous issues.1
  • Training adjustments were made by removing goblin-affine reward signals and filtering data.1
Article not found
CuriousCats.ai

Article

Source Citations