Arena Leaderboards
Crowdsourced head-to-head model evaluations and public leaderboards.
Crowdsourced model leaderboards from the Chatbot Arena team
LMArena, spun out of UC Berkeley's Chatbot Arena project, runs crowdsourced head-to-head AI model evaluations and leaderboards used across the industry. Evals-focused rather than RL environments, but a core evaluation vendor to labs.
Chatbot Arena leaderboard is a de facto industry standard
This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.
| Website | lmarena.ai |
|---|---|
| Domains | Multi-Domain, Machine Learning |
| Location | San Francisco, Berkeley |
| Team size | 26-50 |
| Founded | 2024 |
| Funding | $100M seed (a16z, UC Investments, 2025); $150M at $1.7B valuation (Jan 2026) |
| Founders | Anastasios Angelopoulos @ml_angelopoulos, Wei-Lin Chiang @infwinston, Ion Stoica |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Unknown. No source-backed catalog capability record is available yet.
Products and services supported by official company materials. Reviewed 2026-08-16.
Crowdsourced head-to-head model evaluations and public leaderboards.
Verified public artifacts and documented private commercial inventory.
Structured sources have not yet been migrated for this profile.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Multi-Domain, Machine Learning.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | Pure-play commercial 1 artifacts | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | Unknown |
| Applied Compute Ex-OpenAI trio applying RL to build specialist enterprise models | Enterprise, Machine Learning, Custom Environments | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $80M total at $700M valuation (Oct 2025) | Yes |
| Artificial Analysis Independent benchmarking of AI models across intelligence, speed, and price | Multi-Domain, Machine Learning | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $2.6M (2024) | Unknown |
| Bespoke Labs Data curation and RL environment recipes from ex-Google DeepMind researchers | Coding, Machine Learning | Pure-play commercial Catalog unknown | Mountain View, Menlo Park, Bangalore, San Francisco | 11-25 | ~$40M (reported) | Unknown |
| Collinear Enterprise simulation, judges, and long-horizon trajectory generation | Enterprise, Long Horizon, Machine Learning, Simulation | Pure-play commercial Catalog unknown | Mountain View, Sunnyvale | 11-25 | - | Unknown |