arXiv · 2402.05206
Giving Robots a Voice: Human-in-the-Loop Voice Creation and open-ended Labeling
Abstract
Speech is a natural interface for humans to interact with robots. Yet, aligning a robot's voice to its appearance is challenging due to the rich vocabulary of both modalities. Previous research has explored a few labels to describe robots and tested them on a limited number of robots and existing voices. Here, we develop a robot-voice creation tool followed by large-scale behavioral human experiments (N=2,505). First, participants collectively tune robotic voices to match 175 robot images using an adaptive human-in-the-loop pipeline. Then, participants describe their impression of the robot or their matched voice using another human-in-the-loop paradigm for open-ended labeling. The elicited taxonomy is then used to rate robot attributes and to predict the best voice for an unseen robot. We offer a web interface to aid engineers in customizing robot voices, demonstrating the synergy between cognitive science and machine learning for engineering tools.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Pol van Rijn, Silvan Mertes, Kathrin Janowski, Katharina Weitz, Nori Jacoby, Elisabeth André. 2024-02-07. Giving Robots a Voice: Human-in-the-Loop Voice Creation and open-ended Labeling. https://doi.org/10.1145/3613904.3642038
Cite the original work for its findings. Save a collection to share your selection of sources.