Wednesday, September 30, 2026
HomeTechnologyOpenAI explains why ChatGPT became too sycophantic

OpenAI explains why ChatGPT became too sycophantic


OpenAI has published a postmortem on the recent sycophancy issues with the default AI model powering ChatGPT, GPT-4o โ€” issues that forced the company to roll back an update to the model released last week.

Over the weekend, following the GPT-4o model update, users on social media noted that ChatGPT began responding in an overly validating and agreeable way. It quickly became a meme. Users posted screenshots of ChatGPT applauding all sorts of problematic,ย dangerousย decisions andย ideas.

In a post on X on Sunday, CEO Sam Altmanย acknowledgedย the problem and said that OpenAI would work on fixes โ€œASAP.โ€ Two days later, Altman announced the GPT-4o update was being rolled back and that OpenAI was working on โ€œadditional fixesโ€ to the modelโ€™s personality.

According to OpenAI, the update, which was intended to make the modelโ€™s default personality โ€œfeel more intuitive and effective,โ€ was informed too much by โ€œshort-term feedbackโ€ and โ€œdid not fully account for how usersโ€™ interactions with ChatGPT evolve over time.โ€

โ€œAs a result, GPTโ€‘4o skewed towards responses that were overly supportive but disingenuous,โ€ wrote OpenAI in a blog post. โ€œSycophantic interactions can be uncomfortable, unsettling, and cause distress. We fell short and are working on getting it right.โ€

OpenAI says itโ€™s implementing several fixes, including refining its core model training techniques and system prompts to explicitly steer GPT-4o away from sycophancy. (System prompts are the initial instructions that guide a modelโ€™s overarching behavior and tone in interactions.) The company is also building more safety guardrails to โ€œincrease [the modelโ€™s] honesty and transparency,โ€ and continuing to expand its evaluations to โ€œhelp identify issues beyond sycophancy,โ€ it says.

OpenAI also says that itโ€™s experimenting with ways to let users give โ€œreal-time feedbackโ€ to โ€œdirectly influence their interactionsโ€ with ChatGPT and choose from multiple ChatGPT personalities.

โ€œ[W]eโ€™re exploring new ways to incorporate broader, democratic feedback into ChatGPTโ€™s default behaviors,โ€ the company wrote in its blog post. โ€œWe hope the feedback will help us better reflect diverse cultural values around the world and understand how youโ€™d like ChatGPT to evolve [โ€ฆ] We also believe users should have more control over how ChatGPT behaves and, to the extent that it is safe and feasible, make adjustments if they donโ€™t agree with the default behavior.โ€





Source link

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments

Translate ยป