2026
Instance-wise Adaptive Scheduling via Derivative-Free Meta-Learning
ICLR 2026poster
Deep Reinforcement Learning has achieved remarkable progress in solving NP-hard scheduling problems. However, existing methods primarily focus on optimizing average performance over training instances, overlooking the core objective of solving each individual instance with high quality. While severa…