KellyBench
An environment for evaluating sequential decision-making under long-horizon, non-stationary conditions.
Open reasoning data and reward models from the ex-Meta AI reasoning lead
General Reasoning, co-founded by Ross Taylor (former Meta AI reasoning/Galactica lead), builds open datasets, reward models, and long-horizon RL infrastructure, with finance-domain environments. It runs the OpenReward project.
OpenReward; founder co-created Galactica and Llama reasoning work at Meta
This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.
| Website | gr.inc |
|---|---|
| Domains | Finance, Long Horizon, Machine Learning |
| Location | London, San Francisco |
| Team size | 1-10 |
| Funding | ~$10.9M (reported) |
| Founders | Ross Taylor @rosstaylor90, Chengxi Taylor, Kip Parker, Thomas Grady |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Unknown. No source-backed catalog capability record is available yet.
Products and services supported by official company materials. Reviewed 2026-08-16.
An environment for evaluating sequential decision-making under long-horizon, non-stationary conditions.
Verified public artifacts and documented private commercial inventory.
Structured sources have not yet been migrated for this profile.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Finance, Long Horizon, Machine Learning.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| AfterQuery Expert human data and RL environments across code, finance, and computer use | Multi-Domain, Coding, Finance | Pure-play commercial 1 artifacts | San Francisco, New York, Seattle | 51-100 | $30.5M total (reported) | Unknown |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | Pure-play commercial Catalog unknown | San Francisco | 1-10 | Seed (Y Combinator) | Unknown |
| Andromede Programmatic generation of long-horizon RL environments | Long Horizon | Pure-play commercial Catalog unknown | Lausanne | 1-10 | - | Unknown |
| Anthromind Medical and long-horizon environments and expert data | Medical, Long Horizon, Data Labeling | Data + environments Catalog unknown | San Francisco | 1-10 | - | Unknown |
| Applied Compute Ex-OpenAI trio applying RL to build specialist enterprise models | Enterprise, Machine Learning, Custom Environments | Pure-play commercial Catalog unknown | San Francisco | 11-25 | $80M total at $700M valuation (Oct 2025) | Yes |