Improving Cross-Problem Vehicle Routing with Locally Augmented Preferences and Representation Disentanglement
Multi-task vehicle routing solvers using locally augmented preferences and representation disentanglement to overcome RL reward-scale disparities and preference-optimization stagnation.