Learning Environments
Game-based environments for training model capabilities.
Game-based RL environments and benchmarks
Good Start Labs builds game-like and long-horizon environments for training and benchmarking AI models. Co-founder Alex Duffy created the AI Diplomacy benchmark, which pits frontier models against each other in the strategy game Diplomacy.
AI Diplomacy benchmark
This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.
| Website | goodstartlabs.com |
|---|---|
| Domains | Games, Long Horizon |
| Location | Brooklyn, New York, Toronto |
| Team size | 1-10 |
| Funding | ~$3.6M (reported) |
| Founders | Alex Duffy, Tyler Marques |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Unknown. No source-backed catalog capability record is available yet.
Products and services supported by official company materials. Reviewed 2026-08-16.
Game-based environments for training model capabilities.
Human-generated data for model training.
Game-based model benchmarks and evaluations.
Data-generation programs built around games.
Verified public artifacts and documented private commercial inventory.
Structured sources have not yet been migrated for this profile.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Games, Long Horizon.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | Pure-play commercial Catalog unknown | San Francisco | 1-10 | Seed (Y Combinator) | Unknown |
| Andromede Programmatic generation of long-horizon RL environments | Long Horizon | Pure-play commercial Catalog unknown | Lausanne | 1-10 | - | Unknown |
| Anthromind Medical and long-horizon environments and expert data | Medical, Long Horizon, Data Labeling | Data + environments Catalog unknown | San Francisco | 1-10 | - | Unknown |
| ARIMLABS Security and long-horizon environments for agentic AI | Cybersecurity, Long Horizon | Pure-play commercial Catalog unknown | Warsaw | 11-25 | - | Unknown |
| Collinear Enterprise simulation, judges, and long-horizon trajectory generation | Enterprise, Long Horizon, Machine Learning, Simulation | Pure-play commercial Catalog unknown | Mountain View, Sunnyvale | 11-25 | - | Unknown |