← Search

Mingjun Pan

1 accepted papers

2025

Preference Optimization for Combinatorial Optimization Problems

ICML 2025poster

Reinforcement Learning (RL) has emerged as a powerful tool for neural combinatorial optimization, enabling models to learn heuristics that solve complex problems without requiring expert knowledge. Despite significant progress, existing RL approaches face challenges such as diminishing reward signal…

Cited by 0SourcePDFScholar