Pure-play commercialCatalog available · 1

LMArena

Crowdsourced model leaderboards from the Chatbot Arena team

LMArena, spun out of UC Berkeley's Chatbot Arena project, runs crowdsourced head-to-head AI model evaluations and leaderboards used across the industry. Evals-focused rather than RL environments, but a core evaluation vendor to labs.

Chatbot Arena leaderboard is a de facto industry standard

UnknownConfidence: unknown·Last verified: Unknown·How verification works

This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.

Websitelmarena.ai
DomainsMulti-Domain, Machine Learning
LocationSan Francisco, Berkeley
Team size26-50
Founded2024
Funding$100M seed (a16z, UC Investments, 2025); $150M at $1.7B valuation (Jan 2026)
FoundersAnastasios Angelopoulos @ml_angelopoulos, Wei-Lin Chiang @infwinston, Ion Stoica

Capability coverage

Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.

Focus areas

Legacy classification; source migration pending

Catalog-evidenced technical capabilities

Unknown. No source-backed catalog capability record is available yet.

Capability catalog

Products and services supported by official company materials. Reviewed 2026-08-16.

Suggest an item
Benchmarkcompany reported

Arena Leaderboards

Crowdsourced head-to-head model evaluations and public leaderboards.

Public / Open
Arena

Environments & datasets

Verified public artifacts and documented private commercial inventory.

Add catalog item →
No verified catalog items recorded yet. This does not mean the company has no environments or datasets.

Sources

Structured sources have not yet been migrated for this profile.

Improve this record

Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.

Similar startups

Also listed under Multi-Domain, Machine Learning.

NameDomainsType / catalogLocationTeamFundingRaising
Rise Data Labs
RL environments and tasks across multiple domains, supported by a large expert network
Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments
Data + environments
3 artifacts
New York, United States11-25-Unknown
AfterQuery
Expert human data and RL environments across code, finance, and computer use
Multi-Domain, Coding, Finance
Pure-play commercial
1 artifacts
San Francisco, New York, Seattle51-100$30.5M total (reported)Unknown
Applied Compute
Ex-OpenAI trio applying RL to build specialist enterprise models
Enterprise, Machine Learning, Custom Environments
Pure-play commercial
Catalog unknown
San Francisco11-25$80M total at $700M valuation (Oct 2025)Yes
Artificial Analysis
Independent benchmarking of AI models across intelligence, speed, and price
Multi-Domain, Machine Learning
Pure-play commercial
Catalog unknown
San Francisco11-25$2.6M (2024)Unknown
Bespoke Labs
Data curation and RL environment recipes from ex-Google DeepMind researchers
Coding, Machine Learning
Pure-play commercial
Catalog unknown
Mountain View, Menlo Park, Bangalore, San Francisco11-25~$40M (reported)Unknown
Collinear
Enterprise simulation, judges, and long-horizon trajectory generation
Enterprise, Long Horizon, Machine Learning, Simulation
Pure-play commercial
Catalog unknown
Mountain View, Sunnyvale11-25-Unknown