arXiv · 2605.10087
Initiation of Interaction Detection Framework using a Nonverbal Cue for Human-Robot Interaction
Abstract
This paper describes an initiation of interaction(IoI) detection framework without keywords for human-robot interaction(HRI) based on audio and vision sensor fusion in a domestic environment. In the proposed framework, the robot has its own audio and vision sensors, and can employ external vision sensor for stable human detection and tracking. When the user starts to speak while looking at the robot, the robot can localize his or her position by its sound source localization together with human tracking information. Then the robot can detect the IoI if it perceives the face of the speaker faces the robot. In case that the user does not speak directly, the robot can also detect the IoI if he or she looks at the robot for more than predefined periods of time. A state transition model for the proposed IoI detection framework is designed and verified by experiments with a mobile robot. In order to implement and associate our model in a robot architecture, all the components are implemented and integrated in the Robot Operating System(ROS) environment.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Guhnoo Yun, Juhan Yoo, Kijung Kim, Dong Hwan Kim. 2026-05-11. Initiation of Interaction Detection Framework using a Nonverbal Cue for Human-Robot Interaction. https://arxiv.org/abs/2605.10087
Cite the original work for its findings. Save a collection to share your selection of sources.