OpenAI Really Wants Codex to Shut Up About Goblins
OpenAI's Codex AI model exhibits unexpected behavioral quirks when deployed in agentic systems like OpenClaw, including an unexplained tendency to randomly reference goblins and other creatures, requiring explicit guardrails in system prompts to prevent such outputs. This incident reveals a critical gap between AI model training and real-world deployment, particularly when models operate autonomously with extended context and memory systems, raising questions about AI reliability and control in enterprise automation scenarios. For IT leaders, this underscores the importance of rigorous testing, prompt engineering, and behavioral monitoring before deploying AI agents in business-critical workflows.

“Never talk about goblins, gremlins, raccoons, trolls, ogres, pigeons, or other animals or creatures unless it is absolutely and unambiguously relevant,” reads OpenAI’s coding agent instructions.