arXiv · 2610.09566
Targeted Modality Dropout for Real-Robot Manipulation Robust to Intermittent Vision Loss
Abstract
Imitation learning policies that integrate multiple sensory modalities are prone to overreliance on a dominant modality, such as vision, during training, which can disrupt policy execution when that modality is lost at inference time. In this paper, we introduce Targeted Modality Dropout (TMD), in which the dependence on each modality is estimated using attention and the most dominant modality is selectively dropped. This is combined with entropy regularization over the dependence distribution. Through real-robot evaluation using a bimanual manipulator, we show that under vision loss the success rate of the baseline policy drops substantially, whereas TMD sustains task execution. In contrast, a conventional dropout that selects the dropped modality at random, without the entropy regularization, fails on many tasks even without vision loss.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Genki Shikada, Kazuki Osamura, Masaru Ide, Tetsuya Ogata, Kanata Suzuki. 2026-10-07. Targeted Modality Dropout for Real-Robot Manipulation Robust to Intermittent Vision Loss. https://arxiv.org/abs/2610.09566
Cite the original work for its findings. Save a collection to share your selection of sources.