Sunday, September 27, 2026
HomeTechnologyOpenAI and Anthropic researchers decry 'reckless' safety culture at Elon Musk's xAI

OpenAI and Anthropic researchers decry ‘reckless’ safety culture at Elon Musk’s xAI


AI safety researchers from OpenAI, Anthropic, and other organizations are speaking out publicly against the โ€œrecklessโ€ and โ€œcompletely irresponsibleโ€ safety culture at xAI, the billion-dollar AI startup owned by Elon Musk.

The criticisms follow weeks of scandals at xAI that have overshadowed the companyโ€™s technological advances.

Last week, the companyโ€™s AI chatbot, Grok, spouted antisemitic comments and repeatedly called itself โ€œMechaHitler.โ€ Shortly after xAI took its chatbot offline to address the problem, it launched an increasingly capable frontier AI model, Grok 4, which TechCrunch and others found to consult Elon Muskโ€™s personal politics for help answering hot-button issues. In the latest development, xAI launched AI companions that take the form of a hyper-sexualized anime girl and an overly aggressive panda.

Friendly joshing among employees of competing AI labs is fairly normal, but these researchers seem to be calling for increased attention to xAIโ€™s safety practices, which they claim to be at odds with industry norms.

โ€œI didnโ€™t want to post on Grok safety since I work at a competitor, but itโ€™s not about competition,โ€ said Boaz Barak, a computer science professor currently on leave from Harvard to work on safety research at OpenAI, in a Tuesday post on X. โ€œI appreciate the scientists and engineers at xAI but the way safety was handled is completely irresponsible.โ€

Barak particularly takes issue with xAIโ€™s decision to not publish system cards โ€” industry standard reports that detail training methods and safety evaluations in a good faith effort to share information with the research community. As a result, Barak says itโ€™s unclear what safety training was done on Grok 4.

OpenAI and Google have a spotty reputation themselves when it comes to promptly sharing system cards when unveiling new AI models. OpenAI decided not to publish a system card for GPT-4.1, claiming it was not a frontier model. Meanwhile, Google waited months after unveiling Gemini 2.5 Pro to publish a safety report. However, these companies historically publish safety reports for all frontier AI models before they enter full production.

Techcrunch event

San Francisco
|
October 27-29, 2025

Barak also notes that Grokโ€™s AI companions โ€œtake the worst issues we currently have for emotional dependencies and tries to amplify them.โ€ In recent years, weโ€™ve seen countless stories of unstable people developing concerning relationship with chatbots, and how AIโ€™s over-agreeable answers can tip them over the edge of sanity.

Samuel Marks, an AI safety researcher with Anthropic, also took issue with xAIโ€™s decision not to publish a safety report, calling the move โ€œreckless.โ€

โ€œAnthropic, OpenAI, and Googleโ€™s release practices have issues,โ€ Marks wrote in a post on X. โ€œBut they at least do something, anything to assess safety pre-deployment and document findings. xAI does not.โ€

The reality is that we donโ€™t really know what xAI did to test Grok 4. In a widely shared post in the online forum LessWrong, one anonymous researcher claims that Grok 4 has no meaningful safety guardrails based on their testing.

Whether thatโ€™s true or not, the world seems to be finding out about Grokโ€™s shortcomings in real time. Several of xAIโ€™s safety issues have since gone viral, and the company claims to have addressed them with tweaks to Grokโ€™s system prompt.

OpenAI, Anthropic, and xAI did not respond to TechCrunch request for comment.

Dan Hendrycks, a safety adviser for xAI and director of the Center for AI Safety, posted on X that the company did โ€œdangerous capability evaluationsโ€ on Grok 4, indicating that the company did some pre-deployment testing for safety concerns. However, the results to those evaluations have not been publicly shared.

โ€œIt concerns me when standard safety practices arenโ€™t upheld across the AI industry, like publishing the results of dangerous capability evaluations,โ€ said Steven Adler, an independent AI researcher who previously led dangerous capability evaluations at OpenAI, in a statement to TechCrunch. โ€œGovernments and the public deserve to know how AI companies are handling the risks of the very powerful systems they say theyโ€™re building.โ€

Whatโ€™s interesting about xAIโ€™s questionable safety practices is that Musk has long been one of the AI safety industryโ€™s most notable advocates. The billionaire owner of xAI, Tesla, and SpaceX has warned many times about the potential for advanced AI systems to cause catastrophic outcomes for humans, and heโ€™s praised an open approach to developing AI models.

And yet, AI researchers at competing labs claim xAI is veering from industry norms around safely releasing AI models. In doing so, Muskโ€™s startup may be inadvertently making a strong case for state and federal lawmakers to set rules around publishing AI safety reports.

There are several attempts at the state level to do so. California state Sen. Scott Wiener is pushing a bill that would require leading AI labs โ€” likely including xAI โ€” to publish safety reports, while New York Gov. Kathy Hochul is currently considering a similar bill. Advocates of these bills note that most AI labs publish this type of information anyway โ€” but evidently, not all of them do it consistently.

AI models today have yet to exhibit real-world scenarios in which they create truly catastrophic harms, such as the death of people or billions of dollars in damages. However, many AI researchers say that this could be a problem in the near future given the rapid progress of AI models, and the billions of dollars Silicon Valley is investing to further improve AI.

But even for skeptics of such catastrophic scenarios, thereโ€™s a strong case to suggest that Grokโ€™s misbehavior makes the products it powers today significantly worse.

Grok spread antisemitism around the X platform this week, just a few weeks after the chatbot repeatedly brought up โ€œwhite genocideโ€ in conversations with users. Soon, Musk has indicated that Grok will be more ingrained in Tesla vehicles, and xAI is trying to sell its AI models to The Pentagon and other enterprises. Itโ€™s hard to imagine that people driving Muskโ€™s cars, federal workers protecting the U.S., or enterprise employees automating tasks will be any more receptive to these misbehaviors than users on X.

Several researchers argue that AI safety and alignment testing not only ensures that the worst outcomes donโ€™t happen, but they also protect against near-term behavioral issues.

At the very least, Grokโ€™s incidents tend to overshadow xAIโ€™s rapid progress in developing frontier AI models that best OpenAI and Googleโ€™s technology, just a couple years after the startup was founded.





Source link

RELATED ARTICLES

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Most Popular

Recent Comments

Translate ยป