Bregman Centroid Guided Cross-Entropy Method
Abstract
BC-EvoCEM improves ensemble cross-entropy trajectory optimization by using Bregman centroids to aggregate worker outputs and maintain diversity, boosting convergence and solution quality with minimal overhead.
The Cross-Entropy Method (CEM) is a widely adopted trajectory optimizer in model-based reinforcement learning (MBRL), but its unimodal sampling strategy often leads to premature convergence in multimodal landscapes. In this work, we propose Bregman Centroid Guided CEM (BC-EvoCEM), a lightweight enhancement to ensemble CEM that leverages Bregman centroids for principled information aggregation and diversity control. $mathcal{BC-EvoCEM} computes a performance-weighted Bregman centroid across CEM workers and updates the least contributing ones by sampling within a trust region around the centroid. Leveraging the duality between Bregman divergences and exponential family distributions, we show that mathcal{BC-EvoCEM} integrates seamlessly into standard CEM pipelines with negligible overhead. Empirical results on synthetic benchmarks, a cluttered navigation task, and full MBRL pipelines demonstrate that mathcal{BC-EvoCEM}$ enhances both convergence and solution quality, providing a simple yet effective upgrade for CEM.
Get this paper in your agent:
hf papers read 2506.02205 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper