

- Published on 23 Jun 2025
- Last updated on 2 Mar 2026
- Reading Time: 8 minutes
Hello from LMArena: The Community Platform for Exploring Frontier AI
At LMArena, everything starts with the community. There have been a lot of new members joining us in the past few months so we thought it would be a good time to reintroduce ourselves!
Created by researchers from UC Berkeley’s SkyLab, LMArena is an open platform where everyone can easily access, explore and interact with the world’s leading AI models. When we launched the first leaderboard two years ago, we weren’t trying to build a company. Instead our interests have always been grounded in research values. From the start, we wanted to know how we might create a rigorous, reproducible, community-led framework for real-world model evaluation.
Since then, the community has helped us evaluate over 400 models across text, vision, coding, and more, casting tens of millions of head-to-head battles that directly shape which AIs rise to the top. Your preferences have already changed how models are trained, which ones get released, and what improvements labs prioritize next.
We believe human preferences are essential to building better AI. They’re subjective, diverse, and complex. Which is why everyone should have the ability to contribute to AI progress.
How LMArena Works
On LMArena, anyone can explore and interact with the world’s leading AI models. The arena is designed like a tournament where models are compared anonymously side by side and users vote for the better response. This structure of anonymous battles, dynamic prompts, and rotating users, was designed to reduce bias and reflect diverse, real-world use cases, making it possible to generate statistically meaningful insights.
With votes, the community helps shape a public leaderboard, making AI progress more transparent, accessible, and grounded in real-world usage. The more people prompt and vote, the more everyone learns together and impacts AI progress. Sometimes you may find two responses to be a tie, or that both responses aren’t up to your standards. You decide. Everyone votes differently, and that’s okay! There’s no pressure to vote if you’re unsure which answer is better. Since there's no payment or external incentive, votes come from intrinsic motivation. That’s what makes this community unique. The votes are high-quality because they're grounded in genuine interest. We have a diverse range of subject-matter experts: people you can’t hire through labeling firms, contributing authentic, thoughtful evaluations on their own prompts.
Over time, these votes in battle mode add up to a public leaderboard that reflects collective, real-world judgment. It’s not based on benchmarks or automated scoring, it’s based on how people actually use AI. In fact, around 70% of prompts each month are fresh, meaning it's impossible for any AI model to predict and plan for what they will be evaluated on, no matter how many variants or evaluations are run. In addition to Battle mode, you can also explore specific models in Side-by-Side chat or Direct chat. In these modes, you can get deeper with the models you prefer. All prompts and any votes in these modes are collected for transparent research, but do not contribute to any leaderboards as the models are not anonymous. We also never share any personal information, only the prompts and votes are collected for research purposes and may be posted publicly.









