RL environment market intelligence
Find the companies - and the environments they actually offer.
A research-backed index of RL environment vendors, benchmarks, datasets, verifiers, and environment platforms for AI labs, buyers, researchers, and investors.
67
companies indexed
21
cataloged artifacts
26
domains
3
raising statuses reported
Database-derived · as of 2026-08-16 · raising status may be company-reported and is not investment advice.
New: artifact-level market intelligence
Public samples, private commercial inventory, access conditions, verifier details, provenance, and machine-readable records.
67 companies
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Handshake Career network turned human-data and RL environments provider via Handshake AI | Multi-Domain, Data Labeling, RLHF | Data + environments 2 artifacts | San Francisco, New York, Bangalore, Berlin | 250+ | Series F, $200M (2022, ~$3.5B valuation) | Unknown |
| Mercor Expert marketplace powering evals and RL environments for frontier labs | Multi-Domain, Data Labeling, RLHF | Acquired / inactive 1 artifacts | San Francisco | 250+ | Series C, $350M at $10B valuation (Oct 2025) | Yes |
| Scale Data-labeling incumbent extending into agent evals and RL environments | Multi-Domain, Coding, Data Labeling, RLHF | Data + environments 2 artifacts | San Francisco, New York, Washington DC, London | 250+ | $1.6B+ raised; Meta invested $14.3B at ~$29B valuation (June 2025) | Unknown |
| Turing AGI infrastructure: coding data and RL environments at scale | Coding, Multi-Domain, Data Labeling, RLHF | Data + environments Catalog unknown | San Francisco, Palo Alto, Gurugram | 250+ | Series E, $111M at $2.2B valuation (2025) | Unknown |
| Micro1 Vetted domain experts for AI training data and evals | Data Labeling, Multi-Domain, RLHF | Data + environments Catalog unknown | Los Angeles | 101-250 | Series A, $35M at $500M valuation (Sept 2025) | Unknown |
| Snorkel Programmatic data platform expanding into expert evals and RL environments | Multi-Domain, Coding, Machine Learning, Data Labeling, RLHF | Data + environments Catalog unknown | San Francisco, Redwood City, New York | 101-250 | Series D, $100M at $1.3B valuation (2025) | Unknown |
| Surge Bootstrapped human-data leader with a dedicated RL environments org | Multi-Domain, Data Labeling, RLHF | Data + environments 11 artifacts | San Francisco, New York, Seattle | 101-250 | Bootstrapped; reported in talks to raise ~$1B at $25B+ valuation (2025) | Yes |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | Pure-play commercial 1 artifacts | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | Unknown |
| Mechanize RL environments to automate software engineering, founded by ex-Epoch AI researchers | Coding | Pure-play commercial 1 artifacts | San Francisco | 51-100 | ~$9.1M (reported) | Unknown |
| Datacurve Frontier coding data and repository RL environments via the Shipd bounty platform | Coding, RLHF | Data + environments Catalog unknown | San Francisco | 26-50 | Series A, $15M led by Chemistry (Oct 2025); $17.7M total | Unknown |
| Deeptune Code and computer-use environments; acquired by Mercor | Coding, Computer Use | Acquired / inactive Catalog unknown | New York | 26-50 | Series A, $43M; acquired by Mercor (2026) | Unknown |
| Duality AI Falcon: reality-grade digital twin simulation for AI and robotics | Robotics, Simulation | Simulator Catalog unknown | San Mateo | 26-50 | - | Unknown |
| Genesis AI Physics simulation engine and foundation model for robotics | Robotics, Simulation, Machine Learning | Simulator Catalog unknown | San Francisco, Paris | 26-50 | Seed, $105M (Eclipse Ventures, Khosla Ventures, July 2025) | Unknown |
| Halluminate Sandboxed RL environments for finance and enterprise workflows | Finance, Enterprise, Browser | Pure-play commercial Catalog unknown | San Francisco | 26-50 | - | Unknown |
| Huzzle Labs Long-horizon code, tool-use, and enterprise workflow environments | Long Horizon, Coding, Enterprise | Pure-play commercial Catalog unknown | London, Berlin, San Francisco | 26-50 | ~$6M (reported) | Unknown |
| LMArena Crowdsourced model leaderboards from the Chatbot Arena team | Multi-Domain, Machine Learning | Pure-play commercial Catalog unknown | San Francisco, Berkeley | 26-50 | $100M seed (a16z, UC Investments, 2025); $150M at $1.7B valuation (Jan 2026) | Unknown |
| Pareto Expert data workforce for RLHF, evals, and tool-use environments | Multi-Domain, Tool Use, Data Labeling, RLHF | Data + environments Catalog unknown | San Francisco | 26-50 | - | Unknown |
| Prime Intellect Open superintelligence stack: compute, RL environments hub, and sandboxes | Machine Learning, Environment Platforms, Multi-Domain | Environment platform Catalog unknown | San Francisco | 26-50 | Series A, $130M at $1B valuation (2026); $150M+ total | Unknown |
| Proximal Long-horizon coding RL environments built from real codebases | Coding, Long Horizon | Pure-play commercial Catalog unknown | San Francisco, Bangalore | 26-50 | - | Unknown |
| Taste Labs Design-domain environments and evals for AI models | Design | Pure-play commercial Catalog unknown | New York | 26-50 | - | Unknown |
| Applied Compute Ex-OpenAI trio applying RL to build specialist enterprise models | Enterprise, Machine Learning, Custom Environments | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $80M total at $700M valuation (Oct 2025) | Yes |
| ARIMLABS Security and long-horizon environments for agentic AI | Cybersecurity, Long Horizon | Pure-play commercial Catalog unknown | Warsaw | 11-25 | - | Unknown |
| Artificial Analysis Independent benchmarking of AI models across intelligence, speed, and price | Multi-Domain, Machine Learning | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $2.6M (2024) | Unknown |
| Bespoke Labs Data curation and RL environment recipes from ex-Google DeepMind researchers | Coding, Machine Learning | Pure-play commercial Catalog unknown | Mountain View, Menlo Park, Bangalore, San Francisco | 11-25 | ~$40M (reported) | Unknown |
| Chakra Labs Dojo: a hub of computer-use and tool-use environments | Computer Use, Tool Use | Pure-play commercial Catalog unknown | Brooklyn | 11-25 | ~$10.1M (reported) | Unknown |
| Collinear Enterprise simulation, judges, and long-horizon trajectory generation | Enterprise, Long Horizon, Machine Learning, Simulation | Pure-play commercial Catalog unknown | Mountain View, Sunnyvale | 11-25 | - | Unknown |
| Epoch AI Nonprofit research institute behind FrontierMath and AI capability benchmarks | Math, Machine Learning | Pure-play commercial Catalog unknown | Remote | 11-25 | Philanthropic grants (nonprofit) | Unknown |
| Gray Swan AI Adversarial red-teaming arenas and safety evals for frontier models | Cybersecurity, Alignment | Pure-play commercial Catalog unknown | Pittsburgh | 11-25 | ~$40M (reported) | Unknown |
| HUD Evals and RL environments platform for computer-use agents | Computer Use, Coding, Long Horizon, Environment Platforms, Custom Environments | Environment platform Catalog unknown | San Francisco, Singapore | 11-25 | $15M raised (YC W25, Exceptional Capital) | Unknown |
| Idler Code environments with realistic execution constraints | Coding | Pure-play commercial Catalog unknown | San Francisco | 11-25 | - | Unknown |
| Incalmo Offensive-security environments for AI agents | Cybersecurity | Pure-play commercial Catalog unknown | San Mateo | 11-25 | - | Unknown |
| Latch Biology data infrastructure turned life-science environments and benchmarks | Science | Pure-play commercial Catalog unknown | San Francisco | 11-25 | Series A, $15M (2022) | Unknown |
| Nous Research Open AI lab behind the Atropos RL environments framework | Machine Learning, Environment Platforms | Environment platform Catalog unknown | New York | 11-25 | Series A, $50M led by Paradigm (~$1B valuation, 2025) | Unknown |
| Preference Model Stealth startup working on preference and reward modeling | Machine Learning, Coding | Pure-play commercial Catalog unknown | San Francisco, Toronto, Seattle | 11-25 | - | Unknown |
| Quesma Security-domain RL environments and binary analysis evals | Cybersecurity, Coding | Pure-play commercial Catalog unknown | Warsaw | 11-25 | - | Unknown |
| ReasonCore Science and code reasoning environments and benchmarks | Science, Coding | Pure-play commercial Catalog unknown | San Francisco | 11-25 | - | Unknown |
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| Akhara Enterprise and code RL environments | Enterprise, Coding | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | Pure-play commercial Catalog unknown | San Francisco | 1-10 | Seed (Y Combinator) | Unknown |
| Andromede Programmatic generation of long-horizon RL environments | Long Horizon | Pure-play commercial Catalog unknown | Lausanne | 1-10 | - | Unknown |
| Anthromind Medical and long-horizon environments and expert data | Medical, Long Horizon, Data Labeling | Data + environments Catalog unknown | San Francisco | 1-10 | - | Unknown |
| BenchFlow Open-source benchmark hub and eval infrastructure for agents | Enterprise, Browser, Coding | Pure-play commercial Catalog unknown | San Francisco | 1-10 | ~$1M (reported) | Unknown |
| Diffuse Labs ML and long-horizon RL environments | Machine Learning, Long Horizon | Pure-play commercial Catalog unknown | Palo Alto, San Francisco | 1-10 | - | Unknown |
| Dissei Finance-domain RL environments | Finance | Pure-play commercial Catalog unknown | London | 1-10 | - | Unknown |
| EdotEnv Long-horizon planning environments for frontier models | Long Horizon, Machine Learning | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Exabite Code RL environments with realistic software execution | Coding | Pure-play commercial Catalog unknown | Remote | 1-10 | - | Unknown |
| General Reasoning Open reasoning data and reward models from the ex-Meta AI reasoning lead | Finance, Long Horizon, Machine Learning | Pure-play commercial Catalog unknown | London, San Francisco | 1-10 | ~$10.9M (reported) | Unknown |
| Good Start Labs Game-based RL environments and benchmarks | Games, Long Horizon | Pure-play commercial Catalog unknown | Brooklyn, New York, Toronto | 1-10 | ~$3.6M (reported) | Unknown |
| Hillclimb Math environments emphasizing verifiable correctness | Math | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Invariant Labs Security testing and analysis for AI agents; acquired by Snyk | Cybersecurity, Alignment | Acquired / inactive Catalog unknown | Zurich | 1-10 | Acquired by Snyk (June 2025) | Unknown |
| Matrices Browser-native training environments for web agents | Browser, Computer Use | Pure-play commercial Catalog unknown | San Francisco | 1-10 | ~$5M (reported) | Unknown |
| Metaphi Code and enterprise RL environments | Coding, Enterprise | Pure-play commercial Catalog unknown | San Francisco, New York | 1-10 | - | Unknown |
| Normal Hardware engineering environments for AI models | Hardware Engineering | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Osmosis Forward-deployed reinforcement learning for AI agents | Machine Learning | Pure-play commercial Catalog unknown | San Francisco | 1-10 | Seed, $7M (CRV, Audacious Ventures, YC) | Unknown |
| Plato High-fidelity replicas of websites and software for agent training | Browser, Enterprise, Simulation | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| pre.dev Software-planning platform offering coding and long-horizon RL environments | Coding, Long Horizon | Pure-play commercial Catalog unknown | Delaware | 1-10 | - | Unknown |
| Refresh Simulation engines with verifiable rewards for coding and computer use | Coding, Computer Use, Simulation | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Scaled Foundations GRID: a simulation-first platform for robot learning | Robotics, Simulation, Machine Learning | Simulator Catalog unknown | Seattle | 1-10 | - | Unknown |
| Sepal AI Science-domain environments; acquired by Mercor | Science | Acquired / inactive Catalog unknown | San Francisco | 1-10 | Acquired by Mercor (Feb 2026) | Unknown |
| SynthLabs Post-training research: synthetic data and scalable RL alignment | Machine Learning, Alignment, RLHF | Data + environments Catalog unknown | San Francisco | 1-10 | Seed (M12 and First Spark Ventures, 2024) | Unknown |
| Tacit Labs Life-science and long-horizon environments for AI models | Science, Long Horizon | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Trajectory Labs Alignment-focused environments for safe agent trajectories | Alignment | Pure-play commercial Catalog unknown | Berkeley, Toronto | 1-10 | - | Unknown |
| Ulam Math RL environments and RLVR trajectories | Math | Pure-play commercial Catalog unknown | Warsaw, London | 1-10 | - | Unknown |
| Vals AI Independent domain-specific benchmarks for legal, finance, and tax AI | Legal, Finance, Medical | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Veris AI High-fidelity simulated environments to train enterprise AI agents | Enterprise, Custom Environments, Simulation | Pure-play commercial Catalog unknown | New York | 1-10 | Seed, $8.5M (Decibel and Acrew, June 2025) | Unknown |
| Vetto AI Code and computer-use environments from ex-DeepMind/Instagram founders | Coding, Computer Use | Pure-play commercial Catalog unknown | San Francisco, São Paulo, London | 1-10 | - | Unknown |
| Vmax Converts proprietary data into RL environments | Machine Learning, Custom Environments | Pure-play commercial Catalog unknown | San Francisco, New York | 1-10 | - | Unknown |