Expert AI Training Data
Expert-created training and evaluation data from Handshake's professional network.
Career network turned human-data and RL environments provider via Handshake AI
Handshake operates the largest early-career network in the US and launched Handshake AI, a human data labs business that recruits PhD-level experts to build evals and RL environments for frontier AI labs. It leverages its network of millions of students and experts for domain-specific data.
Known for its 'Gandalf the Grader' benchmark work
| Website | joinhandshake.com/research/ai |
|---|---|
| Domains | Multi-Domain, Data Labeling, RLHF |
| Location | San Francisco, New York, Bangalore, Berlin |
| Team size | 250+ |
| Founded | 2014 |
| Funding | Series F, $200M (2022, ~$3.5B valuation) |
| Founders | Garrett Lord |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Products and services supported by official company materials. Reviewed 2026-08-16.
Expert-created training and evaluation data from Handshake's professional network.
An open-source reactive agent-as-judge that inspects files and environment state while grading agent work.
A banking-workflow benchmark used to evaluate agents and verifier behavior on artifact-heavy tasks.
Verified public artifacts and documented private commercial inventory.
An open-source reactive agent-as-judge that inspects files and environment state while grading agent work.
A banking-workflow benchmark used to evaluate agents and verifier behavior on artifact-heavy tasks.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Multi-Domain, Data Labeling, RLHF.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | Pure-play commercial 1 artifacts | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | Unknown |
| Anthromind Medical and long-horizon environments and expert data | Medical, Long Horizon, Data Labeling | Data + environments Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Artificial Analysis Independent benchmarking of AI models across intelligence, speed, and price | Multi-Domain, Machine Learning | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $2.6M (2024) | Unknown |
| Datacurve Frontier coding data and repository RL environments via the Shipd bounty platform | Coding, RLHF | Data + environments Catalog unknown | San Francisco | 26-50 | Series A, $15M led by Chemistry (Oct 2025); $17.7M total | Unknown |
| LMArena Crowdsourced model leaderboards from the Chatbot Arena team | Multi-Domain, Machine Learning | Pure-play commercial Catalog unknown | San Francisco, Berkeley | 26-50 | $100M seed (a16z, UC Investments, 2025); $150M at $1.7B valuation (Jan 2026) | Unknown |