AgentDojo
Open evaluation framework with realistic office, banking, travel, and messaging tasks plus adversarial security cases.
Security testing and analysis for AI agents; acquired by Snyk
Invariant Labs, an ETH Zurich spin-off, built tools for testing, tracing, and securing AI agents, discovering attack classes like MCP tool poisoning and rug pulls. It was acquired by Snyk in June 2025 to anchor Snyk's agentic AI security research.
Coined 'tool poisoning' and 'MCP rug pull' attacks
This legacy directory profile is awaiting claim-level source migration. Existing values are retained, not upgraded to verified facts.
| Website | invariantlabs.ai |
|---|---|
| Domains | Cybersecurity, Alignment |
| Location | Zurich |
| Team size | 1-10 |
| Founded | 2024 |
| Funding | Acquired by Snyk (June 2025) |
| Raising | Unknown |
| Founders | Marc Fischer, Luca Beurer-Kellner |
Focus areas and technical capabilities are shown separately. Missing technical evidence remains Unknown.
Unknown. No source-backed catalog capability record is available yet.
Products and services supported by official company materials. Reviewed 2026-08-16.
Open evaluation framework with realistic office, banking, travel, and messaging tasks plus adversarial security cases.
Public registry for exploring agent benchmarks and traces.
Verified public artifacts and documented private commercial inventory.
Structured sources have not yet been migrated for this profile.
Company representatives and researchers can propose sourced changes. Submissions do not directly overwrite editorial data.
Also listed under Cybersecurity, Alignment.
| Name | Domains | Type / catalog | Location | Team | Funding | Raising |
|---|---|---|---|---|---|---|
| Rise Data Labs RL environments and tasks across multiple domains, supported by a large expert network | Computer Use, Coding, Finance, Cybersecurity, Legal, Data Labeling, Enterprise, Multi-Domain, RLHF, Custom Environments | Data + environments 3 artifacts | New York, United States | 11-25 | - | Unknown |
| Andon Labs Long-horizon autonomy benchmarks like Vending-Bench | Long Horizon, Alignment | Pure-play commercial Catalog unknown | San Francisco | 1-10 | Seed (Y Combinator) | Unknown |
| ARIMLABS Security and long-horizon environments for agentic AI | Cybersecurity, Long Horizon | Pure-play commercial Catalog unknown | Warsaw | 11-25 | - | Unknown |
| Gray Swan AI Adversarial red-teaming arenas and safety evals for frontier models | Cybersecurity, Alignment | Pure-play commercial Catalog unknown | Pittsburgh | 11-25 | ~$40M (reported) | Unknown |
| Incalmo Offensive-security environments for AI agents | Cybersecurity | Pure-play commercial Catalog unknown | San Mateo | 11-25 | - | Unknown |
| Quesma Security-domain RL environments and binary analysis evals | Cybersecurity, Coding | Pure-play commercial Catalog unknown | Warsaw | 11-25 | - | Unknown |