Arrow Research search
Back to ICRA

ICRA 2009

Mouth gesture and voice command based robot command interface

Conference Paper Visual Servoing - I Artificial Intelligence · Robotics

Abstract

In this paper we present a voice command and mouth gesture based robot command interface which is capable of controlling three degrees of freedom. The gesture set was designed in order to avoid head rotation and translation, and thus relying solely in mouth movements. Mouth segmentation is performed by using the normalized a* component, as in J. Gomez, et al. , (October 2008). The gesture detection process is carried out by a Gaussian mixture model (GMM) based classifier. After that, a state machine stabilizes the system response by restricting the number of possible movements depending on the initial state. Voice commands are modeled using a hidden Markov model (HMM) isolated word recognition scheme. The interface was designed taking into account the specific pose restrictions found in the DaVinci assisted surgery command console.

Authors

Keywords

  • Mouth
  • Cameras
  • Surges
  • Laparoscopes
  • Surgery
  • Instruments
  • Magnetic heads
  • Hidden Markov models
  • Robot vision systems
  • Arm
  • Speech Recognition
  • Robot Commands
  • Mouth Gestures
  • Degrees Of Freedom
  • Hidden Markov Model
  • State Machine
  • Gaussian Mixture Model
  • Statistical Models
  • Surgeons
  • Head Movements
  • Bounding Box
  • Inactive State
  • Visual Feedback
  • Weak State
  • Video Sequences
  • Segmentation Process
  • Instrument Control
  • Continuous Speech
  • Master Controller
  • Word Error Rate

Context

Venue
IEEE International Conference on Robotics and Automation
Archive span
1984-2025
Indexed papers
30179
Paper id
348443079412542003
v2026.09.13