IncumbentCatalog available · 3

Scale

Data-labeling incumbent extending into agent evals and RL environments

Scale AI is the original data-labeling powerhouse for AI labs and enterprises, now building RL environments and agent evaluation products under its agents and RL environments group. After Meta's 2025 investment and the departure of CEO Alexandr Wang, it lost some lab customers but continues to push into environments.

Publishes SWE-bench Pro benchmark

ConfirmedConfidence: high·Last verified: 2026-08-16·How verification works
Websitescale.com
DomainsMulti-Domain, Coding, Data Labeling, RLHF
LocationSan Francisco, New York, Washington DC, London
Team size250+
Founded2016
Funding$1.6B+ raised; Meta invested $14.3B at ~$29B valuation (June 2025)
FoundersAlexandr Wang @alexandr_wang, Lucy Guo

Capability coverage

Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.

Focus areas

confirmed; see profile sources

Catalog-evidenced technical capabilities

Long horizonCode executionSandboxedHuman in the loopExpert authoredProduction derived

Capability catalog

Products and services supported by official company materials. Reviewed 2026-08-16.

Suggest an item
Data servicecompany reported

Agentic Data and Evaluations

Data and evaluation programs for agentic AI systems.

Commercial
Agentic Solutions

Environments & datasets

Verified public artifacts and documented private commercial inventory.

Add catalog item →
Benchmark

SWE-Bench Pro

Public sample

A long-horizon software-engineering benchmark built from public and proprietary repositories.

Scale · 100% completeConfirmed

Sources

  1. 1. SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks? ↗
    Scale Labs · accessed 2026-08-16
  2. 2. SWE-Bench Pro - Public Dataset ↗
    Scale Labs · accessed 2026-08-16

Improve this record

Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.

Similar startups

Also listed under Multi-Domain, Coding, Data Labeling, RLHF.

NameDomainsType / catalogLocationTeamFundingRaising
Rise Data Labs
RL environments and tasks across multiple domains, supported by a large expert network
Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments
Data + environments
3 artifacts
New York, United States11-25-Unknown
AfterQuery
Expert human data and RL environments across code, finance, and computer use
Multi-Domain, Coding, Finance
Pure-play commercial
1 artifacts
San Francisco, New York, Seattle51-100$30.5M total (reported)Unknown
Akhara
Enterprise and code RL environments
Enterprise, Coding
Pure-play commercial
Catalog unknown
San Francisco1-10-Unknown
Anthromind
Medical and long-horizon environments and expert data
Medical, Long Horizon, Data Labeling
Data + environments
Catalog unknown
San Francisco1-10-Unknown
Artificial Analysis
Independent benchmarking of AI models across intelligence, speed, and price
Multi-Domain, Machine Learning
Pure-play commercial
Catalog unknown
San Francisco11-25$2.6M (2024)Unknown
BenchFlow
Open-source benchmark hub and eval infrastructure for agents
Enterprise, Browser, Coding
Pure-play commercial
Catalog unknown
San Francisco1-10~$1M (reported)Unknown