CASE STUDY 01 / Building now
A common ground for robot models.
Standardized robotics benchmarks on real hardware.
Coop builds standardized, third-party robotics benchmarks so humanoid and manipulation model developers can run their models on the same tasks and get comparable scores.
Results need a shared context.
Robotics labs use different robots, setups, and task definitions. That makes it difficult to understand how model results compare across teams.
Same tasks. Comparable scores.
Coop gives model developers a shared set of tasks on real hardware. Results are organized by robot, model, and task, keeping the protocol, scores, and failures useful as the benchmark grows.
Building the evaluation infrastructure.
My focus is building Coop as a co-founder. Coop is backed by a16z speedrun, SR007. This overview covers our public positioning; customer information, benchmark results, and internal implementation details are not included.