Arrow Research search
Back to IROS

IROS 2012

Control by Gradient Collocation: Applications to optimal obstacle avoidance and minimum torque control

Conference Paper Accepted Paper Artificial Intelligence · Robotics

Abstract

We present a new machine learning algorithm for learning optimal feedback control policies to guide a robot to a goal in the presence of obstacles. Our method works by first reducing the problem of obstacle avoidance to a continuous state, action, and time control problem, and then uses efficient collocation methods to solve for an optimal feedback control policy. This formulation of the obstacle avoidance problem improves over standard approaches, such as potential field methods, by being resistant to local minima, allowing for moving obstacles, handling stochastic systems, and computing feedback control strategies that take into account the robot's (possibly non-linear) dynamics. In addition to contributing a new method for obstacle avoidance, our work contributes to the state-of-the-art in collocation methods for non-linear stochastic optimal control problems in two important ways: (1) we show that taking into account local gradient and second-order derivative information of the optimal value function at the collocation points allows us to exploit knowledge of the derivative information about the system dynamics, and (2) we show that computational savings can be achieved by directly fitting the gradient of the optimal value function rather than the optimal value function itself. We validate our approach on three problems: non-convex obstacle avoidance of a point-mass robot, obstacle avoidance for a 2 degree of freedom robotic manipulator, and optimal control of a non-linear dynamical system.

Authors

Keywords

  • Collision avoidance
  • Equations
  • Robot kinematics
  • Mathematical model
  • Least squares approximation
  • Optimal control
  • Obstacle Avoidance
  • Optimization Problem
  • System Dynamics
  • Value Function
  • Local Minima
  • Continuous-time
  • Nonlinear Systems
  • Optimal Function
  • Second Derivative
  • Feedback Control
  • Control Problem
  • Nonlinear Dynamics
  • Continuous Action
  • Continuous State
  • Optimal Policy
  • Nonlinear Control
  • Optimal Control Problem
  • Robot Manipulator
  • Presence Of Obstacles
  • Hamilton–Jacobi–Bellman
  • Reward Function
  • Reward Rate
  • Objective Function
  • Sum Of Squared Differences
  • Symmetric Positive Definite Matrix
  • Obstacle Position
  • Inverted Pendulum
  • World Coordinate
  • End-effector

Context

Venue
IEEE/RSJ International Conference on Intelligent Robots and Systems
Archive span
1988-2025
Indexed papers
26578
Paper id
153864839380504791
v2026.09.13