Monday, September 28, 2026
HomeTechnologyOpenAI pledges to make changes to prevent future ChatGPT sycophancy

OpenAI pledges to make changes to prevent future ChatGPT sycophancy


OpenAI says itโ€™ll make changes to the way it updates the AI models that power ChatGPT, following an incident that caused the platform to become overly sycophantic for many users.

Last weekend, after OpenAI rolled out a tweakedย GPT-4o โ€” the default model powering ChatGPT โ€” users on social media noted that ChatGPT began responding in an overly validating and agreeable way. It quickly became a meme. Users posted screenshots of ChatGPT applauding all sorts of problematic,ย dangerousย decisionsย andย ideas.

In a post on X last Sunday, CEO Sam Altmanย acknowledgedย the problem and said that OpenAI would work on fixes โ€œASAP.โ€ On Tuesday, Altmanย announcedย the GPT-4o update was being rolled back and that OpenAI was working on โ€œadditional fixesโ€ to the modelโ€™s personality.

The company published a postmortem on Tuesday, and in a blog post Friday, OpenAI expanded on specific adjustments it plans to make to its model deployment process.

OpenAI says it plans to introduce an opt-in โ€œalpha phaseโ€ for some models that would allow certain ChatGPT users to test the models and give feedback prior to launch. The company also says itโ€™ll include explanations of โ€œknown limitationsโ€ for future incremental updates to models in ChatGPT, and adjust its safety review process to formally consider โ€œmodel behavior issuesโ€ like personality, deception, reliability, and hallucination (i.e., when a model makes things up) as โ€œlaunch-blockingโ€ concerns.

โ€œGoing forward, weโ€™ll proactively communicate about the updates weโ€™re making to the models in ChatGPT, whether โ€˜subtleโ€™ or not,โ€ wrote OpenAI in the blog post. โ€œEven if these issues arenโ€™t perfectly quantifiable today, we commit to blocking launches based on proxy measurements or qualitative signals, even when metrics like A/B testing look good.โ€

The pledged fixes come as more people turn to ChatGPT for advice. According to one recent survey by lawsuit financier Express Legal Funding, 60% of U.S. adults have used ChatGPT to seek counsel or information. The growing reliance on ChatGPT โ€” and the platformโ€™s enormous user base โ€” raises the stakes when issues like extreme sycophancy emerge, not to mention hallucinations and other technical shortcomings.

Techcrunch event

Berkeley, CA
|
June 5

BOOK NOW

As one mitigating step, earlier this week, OpenAI said it would experiment with ways to let users give โ€œreal-time feedbackโ€ to โ€œdirectly influence their interactionsโ€ with ChatGPT. The company also said it would refine techniques to steer models away from sycophancy, potentially allow people to choose from multiple model personalities in ChatGPT, build additional safety guardrails, and expand evaluations to help identify issues beyond sycophancy.

โ€œOne of the biggest lessons is fully recognizing how people have started to use ChatGPT for deeply personal advice โ€” something we didnโ€™t see as much even a year ago,โ€ continued OpenAI in its blog post. โ€œAt the time, this wasnโ€™t a primary focus, but as AI and society have co-evolved, itโ€™s become clear that we need to treat this use case with great care. Itโ€™s now going to be a more meaningful part of our safety work.โ€





Source link

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments

Translate ยป