OpenAI has published a postmortem on the recent sycophancy issues with the default AI model powering ChatGPT, GPT-4o โ issues that forced the company to roll back an update to the model released last week.
Over the weekend, following the GPT-4o model update, users on social media noted that ChatGPT began responding in an overly validating and agreeable way. It quickly became a meme. Users posted screenshots of ChatGPT applauding all sorts of problematic,ย dangerousย decisions andย ideas.
In a post on X on Sunday, CEO Sam Altmanย acknowledgedย the problem and said that OpenAI would work on fixes โASAP.โ Two days later, Altman announced the GPT-4o update was being rolled back and that OpenAI was working on โadditional fixesโ to the modelโs personality.
According to OpenAI, the update, which was intended to make the modelโs default personality โfeel more intuitive and effective,โ was informed too much by โshort-term feedbackโ and โdid not fully account for how usersโ interactions with ChatGPT evolve over time.โ
Weโve rolled back last weekโs GPT-4o update in ChatGPT because it was overly flattering and agreeable. You now have access to an earlier version with more balanced behavior.
More on what happened, why it matters, and how weโre addressing sycophancy: https://t.co/LOhOU7i7DC
โ OpenAI (@OpenAI) April 30, 2025
โAs a result, GPTโ4o skewed towards responses that were overly supportive but disingenuous,โ wrote OpenAI in a blog post. โSycophantic interactions can be uncomfortable, unsettling, and cause distress. We fell short and are working on getting it right.โ
OpenAI says itโs implementing several fixes, including refining its core model training techniques and system prompts to explicitly steer GPT-4o away from sycophancy. (System prompts are the initial instructions that guide a modelโs overarching behavior and tone in interactions.) The company is also building more safety guardrails to โincrease [the modelโs] honesty and transparency,โ and continuing to expand its evaluations to โhelp identify issues beyond sycophancy,โ it says.
OpenAI also says that itโs experimenting with ways to let users give โreal-time feedbackโ to โdirectly influence their interactionsโ with ChatGPT and choose from multiple ChatGPT personalities.
โ[W]eโre exploring new ways to incorporate broader, democratic feedback into ChatGPTโs default behaviors,โ the company wrote in its blog post. โWe hope the feedback will help us better reflect diverse cultural values around the world and understand how youโd like ChatGPT to evolve [โฆ] We also believe users should have more control over how ChatGPT behaves and, to the extent that it is safe and feasible, make adjustments if they donโt agree with the default behavior.โ


