Search papers, labs, and topics across Lattice.
This paper introduces LLM-as-Trainer (LaT), a novel training paradigm that leverages a pretrained large language model to enhance multi-task neural solvers for various Vehicle Routing Problem (VRP) variants. By periodically generating a stage-wise guidance vector based on cross-task validation metrics, LaT effectively mitigates biases towards specific VRP variants and improves the overall solution quality. Experimental results demonstrate that LaT significantly outperforms existing state-of-the-art multi-task solvers across both trained and unseen VRP variants, showcasing its effectiveness and generalizability.
Using a large language model as an external trainer, LaT boosts multi-task neural solvers' performance on diverse Vehicle Routing Problems without the computational burden of traditional meta-learning.
Multi-task neural solvers aim to handle multiple Vehicle Routing Problem (VRP) variants within a unified model, avoiding separate training for each constraint combination. However, VRP variants differ in optimization difficulty, while existing methods lack stage-wise feedback on their training status, making the model biased to some specific variants. Although meta-learning can support adaptive training, it typically requires bi-level optimization and additional gradient updates, increasing computational cost. To address this limitation, we propose LLM-as-Trainer (LaT), a plug-and-play training paradigm that uses a pretrained large language model as an external trainer. LaT periodically analyzes cross-task validation metrics to generate a stage-wise guidance vector. This vector is combined with the current task's constraint vector and injected into each encoder layer, providing the neural solver with additional training information during subsequent policy optimization. Experiments on 16 VRP variants show that LaT improves the solution quality of several state-of-the-art multi-task neural solvers on both trained and unseen variants, supporting the effectiveness and generality of the proposed training paradigm.