Equivariant Action Sampling for Reinforcement Learning and Planning

Equivariant Action Sampling for Reinforcement Learning and Planning

Reinforcement learning (RL) algorithms for continuous control tasks require accurate sampling-based action selection. Many tasks, such as robotic manipulation, contain inherent problem symmetries. However, correctly incorporating symmetry into sampling-based approaches remains a challenge. This work addresses the challenge of preserving symmetry in sampling-based planning and control, a key component for enhancing decision-making efficiency in RL. We introduce an action sampling approach that enforces the desired symmetry. We apply our proposed method to a coordinate regression problem and show that the symmetry aware sampling method drastically outperforms the naive sampling approach. We furthermore develop a general framework for sampling-based model-based planning with Model Predictive Path Integral (MPPI). We compare our MPPI approach with standard sampling methods on several continuous control tasks. Empirical demonstrations across multiple continuous control environments validate the effectiveness of our approach, showcasing the importance of symmetry preservation in sampling-based action selection.

Publication:

arXiv e-prints

Pub Date:

December 2024

arXiv:

arXiv:2412.12237

Bibcode:

2024arXiv241212237Z

Keywords:

Computer Science - Robotics;
Computer Science - Artificial Intelligence;
Computer Science - Machine Learning

E-Print:

Published at International Workshop on the Algorithmic Foundations of Robotics (WAFR) 2024. Website: http://lfzhao.com/EquivSampling

ADS

Equivariant Action Sampling for Reinforcement Learning and Planning

Abstract