Reinforcement Learning
Data programs for reinforcement learning.
Frontier coding data and repository RL environments via the Shipd bounty platform
Datacurve (YC W24) supplies high-quality, complex coding data, RLHF traces, and repository-based RL environments to foundation model labs. Its Shipd platform gamifies data collection with bounties completed by 1,400+ vetted software engineers.
DeepSWE benchmark work
This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.
| Website | datacurve.ai |
|---|---|
| Domains | Coding, RLHF |
| Location | San Francisco |
| Team size | 26-50 |
| Founded | 2024 |
| Funding | Series A, $15M led by Chemistry (Oct 2025); $17.7M total |
| Founders | Serena Ge @serenaa_ge, Charley Lee |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Unknown. No source-backed catalog capability record is available yet.
Products and services supported by official company materials. Reviewed 2026-08-16.
Data programs for reinforcement learning.
Long-horizon software-engineering and data-science tasks.
Ready-made commercial datasets.
Benchmark and evaluation data programs.
Trajectory datasets for agent training.
Supervised fine-tuning datasets.
Verified public artifacts and documented private commercial inventory.
Structured sources have not yet been migrated for this profile.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Coding, RLHF.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | Pure-play commercial 1 artifacts | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | Unknown |
| Akhara Enterprise and code RL environments | Enterprise, Coding | Pure-play commercial Catalog unknown | San Francisco | 1-10 | - | Unknown |
| BenchFlow Open-source benchmark hub and eval infrastructure for agents | Enterprise, Browser, Coding | Pure-play commercial Catalog unknown | San Francisco | 1-10 | ~$1M (reported) | Unknown |
| Bespoke Labs Data curation and RL environment recipes from ex-Google DeepMind researchers | Coding, Machine Learning | Pure-play commercial Catalog unknown | Mountain View, Menlo Park, Bangalore, San Francisco | 11-25 | ~$40M (reported) | Unknown |
| Deeptune Code and computer-use environments; acquired by Mercor | Coding, Computer Use | Acquired / inactive Catalog unknown | New York | 26-50 | Series A, $43M; acquired by Mercor (2026) | Unknown |