benchmarks evals
A Benchmark Put Human-to-Robot Transfer on One Track
H2RBench standardizes four manipulation tasks and human-video comparisons.
Summary
H2RBench standardizes four manipulation tasks and human-video comparisons.
H2RBench provides a shared real-to-simulation protocol for comparing ways to train robots from human video. It includes four manipulation tasks and evaluates representative transfer methods under common settings, finding that methods benefit differently from extra human demonstrations. The benchmark is designed to make those comparisons less confounded by task and supervision differences; it does not establish a single winning method for all robots.
Why it matters
H2RBench standardizes four manipulation tasks and human-video comparisons.
Limits and context
- The benchmark is designed to make those comparisons less confounded by task and supervision differences; it does not establish a single winning method for all robots.
Key claims
H2RBench standardizes four manipulation tasks and human-video comparisons.
Qualification: The benchmark is designed to make those comparisons less confounded by task and supervision differences; it does not establish a single winning method for all robots.
Evidence: source-2026-09-22-020
Sources
- arXiv preprint 2609.24778arXiv · primary research
Corrections
No corrections have been recorded for this story.