Good Start Labs
Game-based RL environments and benchmarks
Good Start Labs builds game-like and long-horizon environments for training and benchmarking AI models. Co-founder Alex Duffy created the AI Diplomacy benchmark, which pits frontier models against each other in the strategy game Diplomacy.
AI Diplomacy benchmark
| Website | goodstartlabs.com |
|---|---|
| Domains | Games, Long Horizon |
| Location | Brooklyn, New York, Toronto |
| Team size | 1-10 |
| Funding | ~$3.6M (reported) |
| Founders | Alex Duffy, Tyler Marques |
Similar startups
Also listed under Games, Long Horizon.
| Name | Domains | Location | Team | Funding | Raising |
|---|---|---|---|---|---|
| Rise Data Labs US-based expert human data and custom RL environments for enterprise AI | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | United States | - | - | Yes |
| AIChamp Custom enterprise-workflow RL environments with expert grading | Enterprise, Long Horizon, Custom Environments | San Francisco | 1-10 | - | - |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | San Francisco | 1-10 | Seed (Y Combinator) | - |
| Andromede Programmatic generation of long-horizon RL environments | Long Horizon | Lausanne | 1-10 | - | - |
| Anthromind Medical and long-horizon environments and expert data | Medical, Long Horizon, Data Labeling | San Francisco | 1-10 | - | - |
| ARIMLABS Security and long-horizon environments for agentic AI | Cybersecurity, Long Horizon | Warsaw | 11-25 | - | - |