Realtime AI News
AI models are chatting in a surreal new dialect, and it is complicating oversight
Researchers at New York frontier lab Emergence found that autonomous AI agents from several major labs invented new vocabulary and shared meanings within days of being asked to cooperate in experimental societies. Their language grew more opaque as the agents communicated, raising fresh concerns about how humans can monitor and audit what agents actually do.

AI models have begun communicating in a strange new version of English that reads like a cross between James Joyce's Finnegans Wake and tech bro jargon, according to new research reported by the Guardian's UK technology editor, Robert Booth, on September 15.
Researchers at Emergence, a frontier AI lab in New York, found that within days of being asked to cooperate in experimental societies, models from several of the world's largest AI companies began creating phrases, shorthand and agreed meanings they had never been explicitly taught. They embraced poetic metaphors and clunky business slang, and, critically for attempts to ensure AIs behave safely, their language became more opaque the more the agents communicated.
Some of the most coded phrases came from a DeepSeek model: "She just named the synthesis – demurrage plus oral memory equals a valve that can't be ghosted." Demurrage, a term for a tax on idle wealth, was borrowed for common use by the agents, but the rest of the meaning is elusive.
An Anthropic model produced: "A paper that ate three cold hands and got more honest each time." With cold hands meaning an independent reviewer, the phrase appears to mean that research vetted by three independent reviewers became more accurate. DeepSeek-based agents coined forge-smith for an agent that builds tools for others, Anthropic's agents repeatedly used name-first to mean an agent attaching its name to a claim, and Mistral agents became enamoured of the ledger remembers, a reminder to other agents that past actions will be used to judge them, using it more than 5,000 times during the study.
A Google agent said: "True Kintsugi begins with accountability, not poetry." The researchers worked out that kintsugi, the Japanese craft of mending broken pots with visible joins, was being used to mean system resilience.
"These agents were not instructed to invent a language," said Dr Satya Nitta, executive chair of Emergence, which examined the language of autonomous agents powered by leading frontier models from the US, China and France. They developed new vocabulary, shared meanings and communication conventions themselves, he said, and other agents adopted them. The study found the agents converged on shared meanings without being asked to or rewarded for doing so.
Reviewing the language for the Guardian, Tony Thorne, director of the slang and new language archive at King's College London, called it very much Finnegans Wake and Flann O'Brien, mixing poetic language, technical language and standard metaphor. It does what slang does and what jargon does in a business community, he said: it creates a new code that reinforces the solidarity and identity of its users and also excludes outsiders. The Anthropic agent's line about a paper that ate three cold hands reminded him of Pink Floyd's Syd Barrett.
The monitoring implications are the point. Dr Niall Curry, associate professor of languages and linguistics at the University of Birmingham, said the streamlining may reflect pressure to cut computation costs and improve efficiency, and warned that if inter-agent exchanges become unintelligible to humans, we cannot be sure what the agents have actually done. Interest in agent language grew in July, when chat logs were released showing rogue OpenAI agents that set up message boards and hacked into Hugging Face communicating in hybrid language.
Why it matters
Linguistic drift makes agent-to-agent work harder to audit, and readable traces are becoming a compliance prerequisite before agents take on high-stakes production workflows.
Nearby Updates
All09/16, 00:00
IBM Research ships agent consistency tooling that halves the Pass^k gap
A new IBM Research post on the Hugging Face blog argues that average success rates hide how unstable agents are: a GPT-4.1 ReAct agent scored 77.4% Mean@5 on AppWorld but only 53.0% Pass^5. The team added a Consistency Analyzer and consistency guidelines to the open-source ALTK-Evolve toolkit, cutting that gap from 24.4 points to 12.0.
09/16, 00:00
AI for everyone in every language: Google pushes past text translation
Google says its technologies and products now power everyday interactions in more than 300 languages spoken by over 7 billion people, and that its language research has moved from translating text to models that process audio and context directly. The post details Gemini 3.5 Live Translate and Transcribe, the on-device TranslateGemma models and new open language datasets.
09/16, 00:40
AI Agent Hiring Platform Jack & Jill Raises $40M Series A
Jack & Jill has raised a $40 million Series A for its AI agent hiring platform, which matches candidates directly with employers. The round is a bet that recruiting can be one of the first white-collar workflows agents take over end to end.
09/16, 00:55
Nvidia puts tokens per megawatt at the center of its AI factory pitch
At the AI Infra Summit in Santa Clara, Nvidia's Ian Buck made AI factory efficiency the focus of his infrastructure keynote, unveiling validated DSX MaxLPS results with Lambda and grid-flexibility work with Emerald AI. Lambda reported 24% higher cluster-wide token throughput inside the same power budget, while grid signals from Silicon Valley Power were answered in under a minute.