Search papers, labs, and topics across Lattice.
This paper introduces PAC-DP, a novel approach to training Diffusion Policies (DPs) for robotic manipulation by integrating a PAC-Bayes generalization bound into the training process. By augmenting the standard denoising loss with a Kullback-Leibler divergence regularizer, PAC-DP enhances control over generalization in finite-data scenarios, particularly in low-data regimes. Experimental results show that PAC-DP significantly improves denoising performance and success rates in complex tasks, establishing it as a robust framework for policy learning in robotics.
PAC-DP achieves remarkable improvements in robotic manipulation tasks, especially under low-data conditions, by leveraging a principled PAC-Bayes framework.
Diffusion Policies (DPs) are able to perform complex manipulation tasks. However, DPs are typically trained by minimizing a denoising objective, which provides limited control over generalization in the finite-data regimes common in robotics. In this letter, we propose PAC-DP, an approach that increases the performance of DPs in robotic manipulation tasks. By modeling the DP as a Bayesian neural network, and defining a PAC-Bayes generalization bound, we derive a novel training objective that augments the standard denoising loss with a Kullback-Leibler divergence regularizer between the posterior and prior parameter distributions. From the theoretical perspective, our approach provides a principled approach to regularize the training of DPs without significantly increasing the training time. From the practical point of view, experimental results demonstrate improved denoising performance, lower variational negative log-likelihood, and higher success rates across multiple robotic manipulation benchmarks. Crucially, the largest improvements are observed in low-data training regimes and complex tasks, establishing PAC-DP as a theoretically grounded framework for robot policy learning.