Overview This dataset contains ALL in the wild conversation crowdsourced from Search Arena between March 18, 2025 and May 8, 2025. It includes 24,069 multi turn conversations with search LLMs across diverse intents, languages, and topics—alongside 12,652 human preference votes. The dataset spans approximately 11,000 users across 136 countries, 13 publicly released models, around 90 languages (including 11\% multilingual prompts), and over 5,000 multi turn sessions. While user interaction patterns with general purpose LLMs and traditional search engines are increasingly well understood, we believe that search LLMs represent a new and understudied interface—blending open ended generation with real time retrieval. This hybrid interaction mode introduces new user behaviors: how questions are posed, how retrieved information is interpreted, and what types of responses are preferred. We release this dataset to support analysis of this emerging paradigm in human–AI interaction, grounded in large scale, real world usage. All users consented the terms of use to share their conversational data with us for research purposes, and all entries have been redacted using Google’s Data Loss Preventi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy