← Search

Soung Chang Liew

2 accepted papers

2026

GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits

AAAI 2026technical

The restless multi-armed bandit (RMAB) framework is a popular model with applications across a wide variety of fields. However, its solution is hindered by the exponentially growing state space (with respect to the number of arms) and the combinatorial action space, making traditional reinforcement

Cited by 0SourcePDFScholar
2026

HiveMind: Contribution-Guided Online Prompt Optimization of LLM Multi-Agent Systems

AAAI 2026technical

Recent advances in LLM-based multi-agent systems have demonstrated remarkable capabilities in complex decision-making scenarios such as financial trading and software engineering. However, evaluating each individual agent’s effectiveness and online optimization of underperforming agents remain open

Cited by 0SourcePDFScholar