← Search

Gongpu Chen

3 accepted papers

2026

GINO-Q: Learning an Asymptotically Optimal Index Policy for Restless Multi-armed Bandits

AAAI 2026technical

The restless multi-armed bandit (RMAB) framework is a popular model with applications across a wide variety of fields. However, its solution is hindered by the exponentially growing state space (with respect to the number of arms) and the combinatorial action space, making traditional reinforcement

Cited by 0SourcePDFScholar
2025

Actions Speak Louder Than Words: Rate-Reward Trade-off in Markov Decision Processes

ICLR 2025poster

The impact of communication on decision-making systems has been extensively studied under the assumption of dedicated communication channels. We instead consider communicating through actions, where the message is embedded into the actions of an agent which interacts with the environment in a Markov…

Cited by 1SourcePDFScholar
2025

LotteryCodec: Searching the Implicit Representation in a Random Network for Low-Complexity Image Compression

ICML 2025spotlight

We introduce and validate the lottery codec hypothesis, which states that untrained subnetworks within randomly initialized networks can serve as synthesis networks for overfitted image compression, achieving rate-distortion (RD) performance comparable to trained networks. This hypothesis leads to a…

Cited by 0SourcePDFScholar