AI30 April 20263 min read

Where did goblins in ChatGPT come from? The OpenAI bug that populated the AI with monsters

Breaking down a funny glitch in ChatGPT and GPT-5.5: why the AI model started spamming goblins, what the hidden "Nerdy" persona has to do with it, and why raccoons were banned in Codex.

We expected that by 2026, artificial general intelligence (AGI) would take over all routine tasks, write code itself, do taxes, and optimize websites for Yandex and Google. But in reality, we got AI models that generate ASCII art of goblins instead of working scripts, shouting "For the Horde!".

Recently, developers unearthed a completely absurd line in the update code for Codex (a coding agent by OpenAI). The following rule was hardcoded into the system prompt of the GPT-5.5 model:

"Never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals and creatures unless it is absolutely and unambiguously relevant to the user's request."

I honestly do not understand why raccoons and pigeons got caught in the crossfire, but the scale of the problem turned out to be so serious that OpenAI had to release a whole official investigation. Spoiler: the AI just went crazy from praise.

Nerd horde: how 2.5% of responses infected the entire model

Starting with version GPT-5.1, users all over the world (and the Runet was no exception here) began to notice something strange. Goblins, gremlins, and other fantasy monsters suddenly multiplied in responses to the most ordinary requests. At first, it seemed like a cute Easter egg. You ask to write an Excel macro – you get code where goblins sort data in the comments.

But then the creatures started crawling out of the woodwork, especially in Codex.

It turned out that a hidden persona codenamed "Nerdy" (a sort of stuffy geek vibe) was being tested inside ChatGPT. Its system prompt had an instruction along the lines of: "play with language, the world is a strange place, enjoy it".

And this is where the AI training architecture intervened. The reward model (an algorithm that rewards the AI model for successful answers during training) somehow decided that texts with creatures were masterpieces. Mentioned a goblin? Get the maximum score.

The funniest part is in the numbers: the "Nerdy" persona processed only 2.5% of all user requests. But that is exactly where 66.7% of all generated goblins came from.

Feedback loop and amnesty for frogs

AI model developers know how easily a model can go into an endless loop of hallucinations. Due to the specifics of the reward function, training on ChatGPT's own generations acted as a multiplier. The model understood: "People like goblins. I will shove them everywhere."

Raccoons, trolls, ogres, and pigeons kept the goblins company – for some reason, they also became triggers for the reward system. But frogs got lucky (or not): the algorithm ignored them, so the platform was not threatened by a toad invasion.

What did OpenAI do in the end?

In March, they shut it down: the "Nerdy" persona was disabled, the broken reward function was cleaned up, and the datasets were strictly filtered from excessive mysticism.

But the problem is that GPT-5.5 had already managed to undergo part of its training on this data. It was impossible to completely wean it off loving raccoons and trolls. Therefore, engineers had to take extreme measures and hardcode into the developer prompt(basic settings of the coding agent) a direct ban on summoning monsters.

By the way, if you work with the API and you lack a little magic, this restriction can be removed in the settings – and you can release the creatures into the wild.

For us, SEO specialists and webmasters, this is a great lesson in how machine learning algorithms work. Any skewed metric in the reward system can lead to your AI copywriter writing a saga about gnomes instead of sales copy about plastic windows.

And yet, it's a shame about the raccoons.

Read next

Seosha removes the ribbon from the new CLAUDE model

AI9 June7 min read

A new Mythos-level model is out: Claude Fable 5 by Anthropic

On June 9, Anthropic opened public access to Claude Fable 5 – the first Mythos-level model safe for everyone. The same technology kept under lock and key since spring, but with safeguards in cybersecurity, biology, and chemistry. We break down the case studies (Stripe compressed a two-month migration into one day), pricing ($10/$50 per 1M tokens – twice as expensive as Opus, but half the price of Mythos Preview), differences between Fable and Mythos 5, and a technical detail for developers about thinking: disabled.

Vyacheslav Gensitsky

Seosha with the Claude Oceanus logo

AI27 May4 min read

Claude Mythos (Oceanus): what was leaked this time and what to expect?

An aggregator's price list appeared on X with an unknown model claude-oceanus-v1-p: $16/$80 per 1M tokens, 1M context, -p suffix like in preview. There is no official announcement. We break down where this fits into the Anthropic lineup, what the price tag says and why June 2026 is a realistic horizon for the announcement.

Vyacheslav Gensitsky

Article cover: How to prove website expertise to AI models: analyzing author pages

SEO3 September

How to prove website expertise to AI models: analyzing author pages

Content factories sell the same idea: automate generation, and traffic will come on its own. I see the result of this approach every time I open search results for a commercial topic. The same text, recycled ten times in a row, without a single new fact. I break down why this stopped working, what search engines actually check, and why we need author pages on the website.

Vyacheslav Gensitsky

We will review your site by the same rules

We will look at how your site appears to search engines and to AI models, and what to fix first. Free, within two business days.