HUD
Evals and RL environments platform for computer-use agents
HUD (YC W25) provides a platform for building agent evals and RL environments: teams wrap real software as agent-callable tools in isolated containers, define tasks and rewards, and run evals/RL at scale. Over 50 businesses and frontier labs use it, with 2,500+ environments built and 1.3M+ task runs.
DoorDash and UiPath among users; hosts hosted versions of benchmarks like OSWorld
| Website | hud.ai |
|---|---|
| Domains | Computer Use, Coding, Long Horizon, Agents Infrastructure, Custom Environments |
| Location | San Francisco, Singapore |
| Team size | 11-25 |
| Funding | $15M raised (YC W25, Exceptional Capital) |
| Founders | Lorenss Martinsons, Jay Ram |
Similar startups
Also listed under Computer Use, Coding, Long Horizon, Agents Infrastructure, Custom Environments.
| Name | Domains | Location | Team | Funding | Raising |
|---|---|---|---|---|---|
| Rise Data Labs US-based expert human data and custom RL environments for enterprise AI | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | United States | - | - | Yes |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | - |
| AIChamp Custom enterprise-workflow RL environments with expert grading | Enterprise, Long Horizon, Custom Environments | San Francisco | 1-10 | - | - |
| Akhara Enterprise and code RL environments | Enterprise, Coding | San Francisco | 1-10 | - | - |
| Anchor Browser Reliable browser automation platform for agentic AI | Browser, Agents Infrastructure | Tel Aviv | 1-10 | Seed, $6M (Blumberg Capital, Gradient Ventures, Nov 2025) | - |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | San Francisco | 1-10 | Seed (Y Combinator) | - |