HUD

Evals and RL environments platform for computer-use agents

HUD (YC W25) provides a platform for building agent evals and RL environments: teams wrap real software as agent-callable tools in isolated containers, define tasks and rewards, and run evals/RL at scale. Over 50 businesses and frontier labs use it, with 2,500+ environments built and 1.3M+ task runs.

DoorDash and UiPath among users; hosts hosted versions of benchmarks like OSWorld

Websitehud.ai
DomainsComputer Use, Coding, Long Horizon, Agents Infrastructure, Custom Environments
LocationSan Francisco, Singapore
Team size11-25
Funding$15M raised (YC W25, Exceptional Capital)
FoundersLorenss Martinsons, Jay Ram

Similar startups

Also listed under Computer Use, Coding, Long Horizon, Agents Infrastructure, Custom Environments.

NameDomainsLocationTeamFundingRaising
Rise Data Labs
US-based expert human data and custom RL environments for enterprise AI
Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom EnvironmentsUnited States--Yes
AfterQuery
Expert human data and RL environments across code, finance, and computer use
Multi-Domain, Coding, FinanceSan Francisco, New York, Seattle51-100$30.5M total (reported)-
AIChamp
Custom enterprise-workflow RL environments with expert grading
Enterprise, Long Horizon, Custom EnvironmentsSan Francisco1-10--
Akhara
Enterprise and code RL environments
Enterprise, CodingSan Francisco1-10--
Anchor Browser
Reliable browser automation platform for agentic AI
Browser, Agents InfrastructureTel Aviv1-10Seed, $6M (Blumberg Capital, Gradient Ventures, Nov 2025)-
Andon Labs
Long-horizon autonomy benchmarks like Vending-Bench
Long Horizon, AlignmentSan Francisco1-10Seed (Y Combinator)-