Coding
Environments that train and evaluate models on real software engineering: resolving issues in production repositories, writing and reviewing pull requests, debugging failing test suites, and shipping features end-to-end. Coding is the most mature RL environment category: benchmarks like SWE-bench proved that verifiable rewards from test suites scale, and frontier labs now buy coding environments by the thousand.
29 startups
| Name | Domains | Location | Team | Funding | Raising |
|---|---|---|---|---|---|
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | - |
| Akhara Enterprise and code RL environments | Enterprise, Coding | San Francisco | 1-10 | - | - |
| BenchFlow Open-source benchmark hub and eval infrastructure for agents | Enterprise, Browser, Coding | San Francisco | 1-10 | ~$1M (reported) | - |
| Bespoke Labs Data curation and RL environment recipes from ex-Google DeepMind researchers | Coding, Machine Learning | Mountain View, Menlo Park, Bangalore, San Francisco | 11-25 | ~$40M (reported) | - |
| Cua Open-source computer-use agent infrastructure and environments | Coding, Computer Use, Agents Infrastructure | San Francisco | 1-10 | Pre-seed/Seed (YC X25) | - |
| Datacurve Frontier coding data and repository RL environments via the Shipd bounty platform | Coding, RLHF | San Francisco | 26-50 | Series A, $15M led by Chemistry (Oct 2025); $17.7M total | - |
| Deeptune Code and computer-use environments; acquired by Mercor | Coding, Computer Use | New York | 26-50 | Series A, $43M; acquired by Mercor (2026) | No |
| E2B Open-source cloud sandboxes for AI agents | Agents Infrastructure, Coding | San Francisco | 26-50 | Series A, $21M (Insight Partners, July 2025); $32M total | - |
| Emulated Code and ML RL environments | Coding, Machine Learning | San Francisco | 1-10 | - | - |
| Exabite Code RL environments with realistic software execution | Coding | Remote | 1-10 | - | - |
| Habitat Code and desktop interaction RL environments | Coding, Computer Use | New York | 1-10 | - | - |
| HUD Evals and RL environments platform for computer-use agents | Computer Use, Coding, Long Horizon, Agents Infrastructure, Custom Environments | San Francisco, Singapore | 11-25 | $15M raised (YC W25, Exceptional Capital) | - |
| Huzzle Labs Long-horizon code, tool-use, and enterprise workflow environments | Long Horizon, Coding, Enterprise | London, Berlin, San Francisco | 26-50 | ~$6M (reported) | - |
| Idler Code environments with realistic execution constraints | Coding | San Francisco | 11-25 | - | - |
| Mechanize RL environments to automate software engineering, founded by ex-Epoch AI researchers | Coding | San Francisco | 51-100 | ~$9.1M (reported) | - |
| Metaphi Code and enterprise RL environments | Coding, Enterprise | San Francisco, New York | 1-10 | - | - |
| Originator Long-horizon computer-use environments | Computer Use, Coding, Long Horizon | London, Paris | 1-10 | - | - |
| pre.dev Software-planning platform offering coding and long-horizon RL environments | Coding, Long Horizon | Delaware | 1-10 | - | - |
| Preference Model Stealth startup working on preference and reward modeling | Machine Learning, Coding | San Francisco, Toronto, Seattle | 11-25 | - | - |
| Proximal Long-horizon coding RL environments built from real codebases | Coding, Long Horizon | San Francisco, Bangalore | 26-50 | - | - |
| Quesma Security-domain RL environments and binary analysis evals | Cybersecurity, Coding | Warsaw | 11-25 | - | - |
| ReasonCore Science and code reasoning environments and benchmarks | Science, Coding | San Francisco | 11-25 | - | - |
| Refresh Simulation engines with verifiable rewards for coding and computer use | Coding, Computer Use, Simulation | San Francisco | 1-10 | - | - |
| Rise Data Labs US-based expert human data and custom RL environments for enterprise AI | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | United States | - | - | Yes |
| Runloop Devboxes for training and evaluating coding agents | Coding, Agents Infrastructure | San Francisco | 1-10 | Seed, ~$7M | - |
| Scale Data-labeling incumbent extending into agent evals and RL environments | Multi-Domain, Coding, Data Labeling, RLHF | San Francisco, New York, Washington DC, London | 250+ | $1.6B+ raised; Meta invested $14.3B at ~$29B valuation (June 2025) | - |
| Snorkel Programmatic data platform expanding into expert evals and RL environments | Multi-Domain, Coding, Machine Learning, Data Labeling, RLHF | San Francisco, Redwood City, New York | 101-250 | Series D, $100M at $1.3B valuation (2025) | - |
| Turing AGI infrastructure: coding data and RL environments at scale | Coding, Multi-Domain, Data Labeling, RLHF | San Francisco, Palo Alto, Gurugram | 250+ | Series E, $111M at $2.2B valuation (2025) | - |
| Vetto AI Code and computer-use environments from ex-DeepMind/Instagram founders | Coding, Computer Use | San Francisco, São Paulo, London | 1-10 | - | - |