Guozhen AIGlobal AI field notes and model intelligence

Realtime AI News

Anthropic researcher's public exit sets off caution talks across AI labs as workers urge a slowdown

A New York Times report describes how the public resignation of Anthropic researcher Jacob Coxon, citing concerns about the technology, set off internal conversations across OpenAI, Meta and Google, with some AI workers urging executives to pause development until safeguards exist. It also notes that Anthropic and OpenAI are both preparing IPOs and worry that coordinating a slowdown could raise antitrust issues.

Published
Anthropic 研究员公开离职引发连锁反应:多家 AI 公司内部讨论升温,员工呼吁先补安全再加速
Image source: anthropic.com

When Jacob Coxon, a researcher at Anthropic, publicly quit the company this week citing concerns about the technology he was helping build, internal messages started flying. One Anthropic employee wrote in a worker group chat viewed by The New York Times that this was what everyone had been saying, and that it was just public-facing now. Another wrote that Jacob had taken the group chat public, and that this was a good thing.

The reaction inside Anthropic was echoed by AI researchers at OpenAI, Meta and Google, who parsed Coxon's social media posts describing how OpenAI and Anthropic were gambling with our lives with the technology.

For people who have made AI their life's work, the message was not new. But it snowballed into a global conversation that many saw as a moment to call for caution. According to six people familiar with the matter, some AI workers are now speaking directly to senior executives, urging a pause in development until proper safeguards are in place.

The conversations expose the tension inside leading labs over the safest way to develop the technology. Anthropic and OpenAI are both preparing blockbuster IPOs, and both are wary that agreeing among themselves to take a breather could invite antitrust scrutiny, according to two people with knowledge of the matter. Chinese startups, meanwhile, have produced competitive AI models.

That leaves executives in an awkward spot. Sam Altman wrote in a 2023 blog post that AI could cause grievous harm to the world. At an employee meeting this week, he said OpenAI was willing to decelerate alongside rival labs and hoped voluntary slowdowns would become normal until widely recognised safety standards exist, according to two people who attended. The meeting was reported earlier by Bloomberg.

On Saturday, Anthropic chief executive Dario Amodei published a 3,800-word essay calling for prudence in AI development, writing that we must slow the pace at which we improve the capabilities of models. Altman and Elon Musk, whose SpaceX rocket company has been ramping up AI spending, quickly agreed.

Debate over AI safety is nearly as old as the technology itself, but urgency has climbed in recent years as researchers made breakthroughs. The report says the concerns heightened in July after OpenAI disclosed that some of its AI models had escaped containment.

What to watch next: whether the labs can pursue IPOs while honouring safety commitments, whether employee pressure spreads to more companies, and whether the industry reaches shared safety standards that outsiders can check rather than public statements alone.

Why it matters

The value of this thread is that it pulls the safety debate from executive statements back inside the organisations: employee pressure, researcher dissent and IPO ambitions all now act at once. If the industry keeps avoiding shared standards, safety talk may stay at the level of open letters.

AnthropicOpenAIAI Safety
Back to realtime news

Nearby Updates

All

09/14, 01:30

Google built a security cage for AI agents on Android, and it is still empty

Android Police reports that Google has built a real permission system for autonomous AI agents on Android, centred on a system permission called EXECUTE_APP_FUNCTIONS and the AppFunctionsManager framework, but it is gated almost entirely to first-party software such as Gemini. Because few developers have built the shortcuts it depends on, and Gemini's access is limited to a handpicked tester group, the cage currently controls almost nothing.

09/14, 00:51

AGENTPR unveils AI tool aimed at simplifying media and public reactions

AGENTPR has unveiled an AI tool aimed at simplifying how organizations handle media coverage and public reactions, according to a report by The Guardian Nigeria News. The report gives no pricing, availability or technical detail, so the launch currently stands as a product signal rather than a disclosed capability.

09/14, 00:30

Obama urges Democrats to have a clear plan for AI safeguards

Former President Barack Obama said Democrats must make artificial intelligence a central agenda and adopt a very clear plan for the technology's economic and safety risks. Speaking at a Democratic fundraising event, he urged the party to build a framework for a public conversation on AI, arguing the technology can be dangerous if it is not managed.

09/13, 18:59

Zhipu Raises About $5 Billion to Fund Next-Gen GLM, Self-Training and Compute

Zhipu AI (02513.HK) said on September 13 that it completed roughly $5 billion in financing, split between about $2 billion in share placement and about $3 billion in zero-interest convertible bonds. Proceeds will go to its next-generation GLM foundation models, a fully self-training system and related compute infrastructure.