Arrow Research search
Back to ICML

ICML 2025

Flow-based Domain Randomization for Learning and Sequencing Robotic Skills

Conference Paper Accept (poster) Artificial Intelligence ยท Machine Learning

Abstract

Domain randomization in reinforcement learning is an established technique for increasing the robustness of control policies learned in simulation. By randomizing properties of the environment during training, the learned policy can be robust to uncertainty along the randomized dimensions. While the environment distribution is typically specified by hand, in this paper we investigate the problem of automatically discovering this sampling distribution via entropy-regularized reward maximization of a neural sampling distribution in the form of a normalizing flow. We show that this architecture is more flexible and results in better robustness than existing approaches to learning simple parameterized sampling distributions. We demonstrate that these policies can be used to learn robust policies for contact-rich assembly tasks. Additionally, we explore how these sampling distributions, in combination with a privileged value function, can be used for out-of-distribution detection in the context of an uncertainty-aware multi-step manipulation planner.

Authors

Keywords

  • Reinforcement Learning
  • Domain Randomization
  • Uncertainty
  • Assembly
  • Planning

Context

Venue
International Conference on Machine Learning
Archive span
1993-2025
Indexed papers
16471
Paper id
834876999790101446
v2026.09.13