Why does a simulation company open-source 100,000 hours of egocentric data?
Because training data is scaling fast and evaluation isn't. @bgxc explains on RoboPapers: ego data is the most scalable way to train, simulation is the only scalable way to test. EgoSuite-Open100k is
Robotics has a data problem, and egocentric data is a compelling way to solve it because egocentric data is (1) scalable — it’s cheap to collect — and (2) easily captures the breadth and diversity of real human tasks, while (3) being true to real physics. Lightwheel recently


