← Search

Yi Mei

3 accepted papers

2026

Lifelong Learning with Behavior Consolidation for Vehicle Routing

ICLR 2026poster

Recent neural solvers have demonstrated promising performance in learning to solve routing problems. However, existing studies are primarily based on one-off training on one or a set of predefined problem distributions and scales, i.e., tasks. When a new task arises, they typically rely on either z…

Cited by 0SourcecodeScholar
2026

ParetoHqD: Fast Offline Multiobjective Alignment of Large Language Models Using Pareto High-Quality Data

AAAI 2026technical

Aligning large language models with multiple human expectations and values is crucial for ensuring that they adequately serve a variety of user needs. To this end, offline multiobjective alignment algorithms such as the Rewards-in-Context algorithm have shown strong performance and efficiency. Howev

Cited by 0SourcePDFScholar