TerraLingua This is a dataset generated by the TerraLingua multi agent system to study the emergence of language, culture, and social structure among LLM powered agents. Agents with personality traits compete for resources, communicate through persistent text artifacts, and form communities over thousands of timesteps. The dataset includes raw simulation logs, full LLM reasoning traces, behavioral annotations generated by an AI Anthropologist, and artifact linguistic complexity metrics. The overview of the TerraLingua system and of the AI Anthropologist is shown in the figure below. Paper: Link ArXiv Code: https://github.com/cognizant ai lab/terralingua Dataset dashboard: https://aianthropology.decisionai.ml/ Dataset Summary Total size : ~4.7 GB Experiments : 40 (8 conditions × 5 repetitions) Agent model : DeepSeek R1 32B Annotation models : Claude Sonnet 4.5 (agent & community annotations, novelty scoring), Claude Haiku 4.5 (artifact phylogeny) Grid : 50×50, up to 3,000 timesteps per run Initial agents per run : 20 (with reproduction) Experimental Conditions Each condition isolates one variable against a core baseline. All conditions are run 5 times (repetitions 1–5). Condition Ke…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy