How does languages work in practice?

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion covers emergent, languages, populations from first principles with code examples. Free lesson at https://engineersofai.com/docs/research/paper-breakdowns/2026-05-29-emergent-languages-in-populations-of-language-model-agents-from-token-efficiency

What is the difference between emergent and populations?

See the full breakdown at https://engineersofai.com/docs/research/paper-breakdowns/2026-05-29-emergent-languages-in-populations-of-language-model-agents-from-token-efficiency

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion

:::info Stub — Full Engineering Breakdown Coming This paper has a linked code implementation and was featured on Hugging Face Papers with 3 upvotes. A full breakdown with production viability rating, implementation notes, and honest limitations is being written. Subscribe to AI Letters → :::


Authors	Stine Lyngsø Beltoft et al.
Year	2026
HF Upvotes	3
arXiv	2605.31170
PDF	Download
Code	https://github.com/aisilab/emergent-languages

Abstract

Monitoring autonomous language model agents currently relies mostly on surface behavior. But what happens when agent populations invent new languages with the goal of avoiding human oversight. Here, we study the emergent languages on Moltbook. For this, we build upon the Moltbook Files dataset and apply a two-stage approach consisting of a rule-based heuristic (about 6000 matches) followed by zero-shot classification (518 kept). The resulting categories include token efficiency (166), new natural languages (106), and oversight evasion (59). We conduct both quantitative and qualitative analyses. Our results show that posts proposing new languages for avoiding oversight are judged by DeepSeek-3.2 as being less aligned than the other categories and that all languages can be learned by other language models in-context merely from a description of the language. Moreover, manually studying exemplary cases reveals surprisingly sophisticated steganographic protocols like embedding hidden messages in natural language. Although we cannot be certain about the extent of autonomy in ideation of these languages, our results add up to the evidence that monitoring surface behavior may soon be insufficient for retaining control over agent populations.

Engineering Breakdown

The Problem

But what happens when agent populations invent new languages with the goal of avoiding human oversight.

The Approach

But what happens when agent populations invent new languages with the goal of avoiding human oversight.

Key Results

Although we cannot be certain about the extent of autonomy in ideation of these languages, our results add up to the evidence that monitoring surface behavior may soon be insufficient for retaining control over agent populations.

Research Areas

This paper contributes to the following areas of AI/ML engineering:

Machine learning
Deep learning
Neural networks
Model optimization
AI systems
Languages

:::tip Subscribe Get weekly breakdowns of papers like this in AI Letters - the newsletter for engineers building production AI systems. :::

Back to Research Lab → · Subscribe to AI Letters →

Abstract​

Engineering Breakdown​

The Problem​

The Approach​

Key Results​

Research Areas​