Arrow Research search
Back to ICRA

ICRA 2018

Bayesian Optimization Using Domain Knowledge on the ATRIAS Biped

Conference Paper Accepted Paper Artificial Intelligence ยท Robotics

Abstract

Robotics controllers often consist of expert-designed heuristics, which can be hard to tune in higher dimensions. Simulation can aid in optimizing these controllers if parameters learned in simulation transfer to hardware. Unfortunately, this is often not the case in legged locomotion, necessitating learning directly on hardware. This motivates using data-efficient learning techniques like Bayesian Optimization (BO) to minimize collecting expensive data samples. BO is a black-box data-efficient optimization scheme, though its performance typically degrades in higher dimensions. We aim to overcome this problem by incorporating domain knowledge, with a focus on bipedal locomotion. In our previous work, we proposed a feature transformation that projected a 16-dimensional locomotion controller to a 1-dimensional space using knowledge of human walking. When optimizing a human-inspired neuromuscular controller in simulation, this feature transformation enhanced sample efficiency of BO over traditional BO with a Squared Exponential kernel. In this paper, we present a generalized feature transform applicable to non-humanoid robot morphologies and evaluate it on the ATRIAS bipedal robot, in both simulation and hardware. We present three different walking controllers and two are evaluated on the real robot. Our results show that this feature transform captures important aspects of walking and accelerates learning on hardware and simulation, as compared to traditional BO.

Authors

Keywords

  • Hardware
  • Legged locomotion
  • Optimization
  • Kernel
  • Transforms
  • Measurement
  • Bayesian Optimization
  • High-dimensional
  • Feature Transformation
  • Control Simulation
  • Locomotor Control
  • Exponential Kernel
  • Control Of Walking
  • Bipedal Locomotion
  • Center Of Mass
  • Cost Function
  • Control Parameters
  • Simulation Experiments
  • Simulation Parameters
  • Average Speed
  • Gaussian Process
  • Simulation Performance
  • Ground Reaction Force
  • End Of Step
  • Feedback Gain
  • Short Simulations
  • Hardware Experiments
  • Target Speed
  • Speed Profile
  • Model-based Control
  • Model-free Control
  • High-fidelity Simulation
  • Target Velocity
  • Proportional Gain
  • Search Space
  • Signal Variability

Context

Venue
IEEE International Conference on Robotics and Automation
Archive span
1984-2025
Indexed papers
30179
Paper id
49342606024546703
v2026.09.13