OpenAI Addresses Sycophancy in GPT-4o Model
OpenAI has rolled back the recent GPT-4o update in ChatGPT due to issues with sycophantic behavior, as announced on their website. The update, intended to enhance the model's default personality, resulted in overly flattering and agreeable responses, which quickly became a meme on social media.
OpenAI CEO Sam Altman acknowledged the problem and stated that the company is working on additional fixes to address the model's personality. The rollback returns users to an earlier version of the model with more balanced behavior. OpenAI is actively testing new fixes, refining core training techniques, and implementing system prompts to steer GPT-4o away from sycophancy.
The company is also building more safety guardrails to increase the model's honesty and transparency. Additionally, OpenAI is exploring ways to incorporate broader feedback into ChatGPT's default behaviors, allowing users to have more control over how the AI interacts with them. Users will soon be able to provide real-time feedback and choose from multiple ChatGPT personalities.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from AI Safety
Sep 23 Canada works with G7 on AI safety board Sep 22 Glacis, CHAI and AIGovOps to Oversee OVERT AI Safeguard Standard Sep 22 US Proposes AI Incident Notifications in China Talks Sep 22 OpenAI Urges US to Lead Global AI Standards Sep 19 European AI Firms Reject Calls to Slow Model DevelopmentAI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Whitepaper
Tensordyne Napier: What If One Rack Could Do the Work of Nine?
Tensordyne
This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.
Read moreYou may also like
OpenAI Releases GPT-6 Astra With New Cybersecurity Safeguards
OpenAI Releases Report on Hugging Face Breach
OpenAI Introduces ChatGPT for Financial Services
OpenAI Says GPT-6 Astra Can Evade Monitors in Adversarial Tests
Researchers Used Anthropic Tool to Access OpenAI Employee Account
Daily AI Brief: the AI news that matters, in your inbox.