Robot Wiki

Data, Hardware & Evaluation

The embodied data bottleneck and the machines themselves, with the measurement crisis behind every benchmark claim.

This domain overview lists 7 articles, ordered as where embodied training data comes from, then the rigs that produce it, then why two reported benchmark numbers are usually not comparable at all.

  1. Robot-hours versus LLM tokens: the log-log reality of embodied data and teleop-farm economics.

  2. Open X-Embodiment, DROID, BridgeData V2, AgiBot World, RoboMIND: five datasets compared.

  3. Arms, humanoids, hands, sensors, and compute: a buyer's guide from SO-101 to Jetson Thor.

  4. ALOHA, GELLO, UMI, and VR teleop: cost, data quality, throughput, and the embodiment gap.

  5. Why N-of-10 trials and unreported variance mislead: 95% per-step success is unusable at 30 steps.

  6. The installed base robot learning is trying to enter, and the jam-rate arithmetic that decides whether a 99 percent cell ships.

  7. Data capture, schemas, training, simulation, evaluation, serving and robot integration as one reproducible system rather than a model checkpoint.