EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration
- URL: http://arxiv.org/abs/2607.19745v2
- Date: Thu, 23 Jul 2026 16:20:02 GMT
- Title: EgoRecovery: Acquiring Failure Recovery Ability Through Human Recovery Demonstration
- Abstract summary: We show that egocentric human data capturing failure recovery processes provides a scalable alternative to robot teleoperation.<n>By efficiently arranging task-level failure configurations and recording short recovery segments, human operators can generate more than 10x as much valid recovery data per hour.
- Score: 75.13534797264485
- License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
- Abstract: Robust embodied robots should be able to recover from failures and retry tasks in order to operate reliably in unstructured and noisy real-world environments. Achieving this capability requires training policies on data that captures recovery behaviors. However, collecting such data through robot teleoperation is difficult to scale, as it is time-consuming to induce diverse failure states, perform corrective actions, and reset the environment. This challenge is further exacerbated by the high diversity of failure modes, which demands substantially more recovery data than success demonstrations. In this work, we show that egocentric human data capturing failure recovery processes provides a scalable alternative. By efficiently arranging task-level failure configurations and recording short recovery segments, human operators can generate more than 10x as much valid recovery data per hour compared to robot teleoperation under our protocol. To address the embodiment gap between human and robot, we propose EgoRecovery, a co-training framework for learning recovery behavior, where human recovery demonstrations are aligned to a compact corrective-intent space shared with robot data, which captures the timing and magnitude of correction. Only a small number of robot recovery demonstrations are required to connect this intent to executable robot actions. At deployment, a learned recovery gate predicts when correction is needed from robot observations and activates the corrective intent only in recovery states. Experiments on real-world recovery tasks show that EgoRecovery improves success from failure starts over robot-only recovery, direct co-training with human recovery data, and direct intent-transfer baselines.
Related papers
- LIBERO-RECOVER: Beyond Task Success Towards Failure Recovery in Robotic Manipulation Models [48.57715789642472]
We introduce LIBERO-Recover Benchmark, a large scale benchmark for failure recovery in robotic manipulation.<n>We collect real execution failures from SOTA embodied models and construct 1,000+ scenarios across four recovery levels.<n>We evaluate four core capabilities: spatial understanding, object structure reasoning, interaction understanding, and topological reasoning.
arXiv Detail & Related papers (2026-09-04T14:15:27Z) - Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data [60.94418315274988]
Ego2Robot supports both curated and in-the-wild videos, producing 18,561 hours of robot training data spanning 15 robot morphologies.<n>Joint pretraining on Ego2Robot-synthesized and robot data consistently improves out-of-distribution generalization across multiple types.
arXiv Detail & Related papers (2026-08-03T17:52:26Z) - DenseReward: Dense Reward Learning via Failure Synthesis for Robotic Manipulation [67.08835970838996]
Reinforcement learning holds great promise for improving robot policies beyond the limits of imitation learning.<n>Two key challenges remain: acquiring diverse failure data at scale and obtaining fine-grained reward signals beyond sparse trajectory-level success labels.<n>We introduce DenseReward, a dense robotic reward model that addresses both challenges.
arXiv Detail & Related papers (2026-07-14T17:59:29Z) - HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation [24.40844636932594]
HELP is a Human-Efficient Large-scale robot Post-training pipeline.<n>Two specialized operators supervise twelve robots concurrently.<n> HELP achieves 80%--95% success rates and improves task throughput by 1.7$times$--4.2$times$ over the base model.
arXiv Detail & Related papers (2026-07-08T03:03:38Z) - HumanScale: Egocentric Human Video Can Outperform Real-Robot Data for Embodied Pretraining [93.5192770515694]
Embodied foundation models are expected to benefit from data scaling like large language models, but face a much tighter data bottleneck.<n>Teleoperated real-robot trajectories remain the dominant pretraining source due to their precise action supervision and embodiment alignment.<n>Embodied foundation models pretrained on egocentric data achieve a 24% lower validation loss on real-robot action prediction.
arXiv Detail & Related papers (2026-06-18T17:37:34Z) - Learning Actionable Manipulation Recovery via Counterfactual Failure Synthesis [21.197844940385725]
Current failure-learning paradigms rely on either costly and unsafe real-world data collection or simulator-based perturbations.<n>We introduce Dream2Fix, a framework that synthesizes photorealistic, counterfactual failure rollouts directly from successful real-world demonstrations.<n>By perturbing actions within a generative world model, Dream2Fix creates paired failure-language data without relying on simulators.
arXiv Detail & Related papers (2026-03-13T19:02:58Z) - EgoHumanoid: Unlocking In-the-Wild Loco-Manipulation with Robot-Free Egocentric Demonstration [67.13034606664333]
EgoHumanoid is the first framework to co-train a vision-language-action policy using egocentric human demonstrations.<n>A portable system for scalable human data collection is developed.
arXiv Detail & Related papers (2026-02-10T18:59:03Z) - Human-in-the-Loop Failure Recovery with Adaptive Task Allocation [2.518621093955008]
We propose an adaptive method for allocating robotic failures to human operators (ARFA)<n>For every failure to be resolved, a reward function calculates expected outcomes based on operator capabilities and historical data, task urgency, and current workload distribution.<n>Our simulations and user studies show that ARFA outperforms random allocation, significantly reducing robot idle time, improving overall system performance, and leading to a more distributed workload among operators.
arXiv Detail & Related papers (2026-02-03T14:55:48Z) - RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction [23.89121398540929]
We introduce RaC, a new phase of training on human-in-the-loop rollouts after imitation learning pre-training.<n>In RaC, we fine-tune a robotic policy on human intervention trajectories that illustrate recovery and correction behaviors.<n>We show that RaC outperforms the prior state-of-the-art using 10$times$ less data collection time and samples.
arXiv Detail & Related papers (2025-09-09T17:41:29Z) - Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition [48.65867987106428]
We introduce a novel system for joint learning between human operators and robots.
It enables human operators to share control of a robot end-effector with a learned assistive agent.
It reduces the need for human adaptation while ensuring the collected data is of sufficient quality for downstream tasks.
arXiv Detail & Related papers (2024-06-29T03:37:29Z) - MIRACLE: Inverse Reinforcement and Curriculum Learning Model for
Human-inspired Mobile Robot Navigation [13.824617183645291]
In emergency scenarios, mobile robots must navigate like humans, interpreting stimuli to locate potential victims rapidly without interfering with first responders.
We propose a solution, MIRACLE, that employs gamified learning to gather stimuli-driven human navigational data.
This data is then used to train a Deep Inverse Maximum Entropy Reinforcement Learning model, reducing reliance on demonstrator abilities.
arXiv Detail & Related papers (2023-12-06T18:13:21Z) - Revisiting the Adversarial Robustness-Accuracy Tradeoff in Robot
Learning [121.9708998627352]
Recent work has shown that, in practical robot learning applications, the effects of adversarial training do not pose a fair trade-off.
This work revisits the robustness-accuracy trade-off in robot learning by analyzing if recent advances in robust training methods and theory can make adversarial training suitable for real-world robot applications.
arXiv Detail & Related papers (2022-04-15T08:12:15Z)
This list is automatically generated from the titles and abstracts of the papers in this site.
This site does not guarantee the quality of this site (including all information) and is not responsible for any consequences.