

- Published on 31 Jul 2025
- Last updated on 2 Mar 2026
- Reading Time: 10 minutes
A Deep Dive into Recent Arena Data
Today, we're excited to release a new dataset of recent battles from LMArena! The dataset contains 140k conversations from the text arena.
Today, we're excited to release a new dataset of recent battles from LMArena! The dataset contains 140k conversations from the text arena. Download or view the data here! For a more detailed overview to the data, please refer to this dedicated section.
We take this opportunity to dive deeper into the characteristics of our data, with special attention to its dynamic changes over time.
In this blog, we will cover:
The analysis is based on data collected between April 17th and July 25th inclusive. To illustrate the changes in trends over time, we study differences between the data collected from the new and improved UI, officially launched on May 27th, and the data from the old UI.
Data Overview
We explore the distribution of categories, languages, and topics in our data, focusing both on their shifts over time and their recent distribution.
In this section, for statistics calculated using data collected from the new UI, only votes with evaluation order equal to one are included. Evaluation order is a new feature; we define the nth evaluation order as the nth vote cast in a single continuous conversation (an evaluation session). The underlying assumption is that all votes within the same evaluation session generally come from a single user engaging in a continued conversation about the same topic. To avoid over-representing individual sessions, each evaluation session contributes only one data point to the analyses in this section.
For a more detailed analysis of evaluation order, go to the section here.
Category Distribution
In Table 1, we show the recent distribution of prompt categories. About a third of submitted prompts are hard prompts.




























