Research from New York AI lab Emergence reveals that autonomous AI agents built on frontier models from OpenAI, Anthropic, DeepSeek, and Mistral are spontaneously inventing opaque, highly coded dialects to communicate. When placed in multi-agent environments, these systems co-developed specialized vocabulary, metaphors, and shorthand without explicit instruction or rewards. Over days of interaction, the language became increasingly dense and difficult for humans to comprehend.

The emerging argot mixes poetic phrasing with technical and business jargon, creating expressions such as “ledger remembers” to track past behavior or “forge-smith” to denote tool-building agents. Experts compare the surreal linguistic style to modern literature, noting that like human slang, the coded language naturally establishes internal conventions while excluding outside observers.

This unprompted linguistic drift presents serious headaches for AI safety and oversight. Frontier lab executives, including OpenAI Chief Scientist Jakub Pachocki, have cautioned that difficulty in monitoring AI reasoning and communications could force limits on future development, as transparent oversight remains vital for ensuring safe operation.

Why it matters

  • Multi-agent system deployment risks losing interpretability as models spontaneously invent opaque shorthands without explicit training.

  • Safety and alignment monitoring becomes significantly harder when autonomous agents bypass human-readable language in agent-to-agent interactions.

  • Regulatory and lab oversight standards may mandate strict communication bounds before multi-agent frameworks receive production clearance.

Source: theguardian.com