← Search

Vivienne Huiling Wang

5 accepted papers

2026

Learning Multi-Timescale Abstractions for Hierarchical Combinatorial Planning

ICML 2026poster

The combination of exponentially large action spaces, stochastic dynamics, and long-horizon decision-making under limited resources makes Sequential Stochastic Combinatorial Optimization (SSCO) particularly challenging for reinforcement learning. Hierarchical Reinforcement Learning (HRL) offers a na…

Cited by 0SourceScholar
2025

Hierarchical Reinforcement Learning with Uncertainty-Guided Diffusional Subgoals

ICML 2025poster

Hierarchical reinforcement learning (HRL) learns to make decisions on multiple levels of temporal abstraction. A key challenge in HRL is that the low-level policy changes over time, making it difficult for the high-level policy to generate effective subgoals. To address this issue, the high-level po…

Cited by 0SourcePDFScholar
2025

Multi-Scale Fusion for Object Representation

ICLR 2025poster

Representing images or videos as object-level feature vectors, rather than pixel-level feature maps, facilitates advanced visual tasks. Object-Centric Learning (OCL) primarily achieves this by reconstructing the input under the guidance of Variational Autoencoder (VAE) intermediate representation to…

2024

Probabilistic Subgoal Representations for Hierarchical Reinforcement Learning

ICML 2024poster

In goal-conditioned hierarchical reinforcement learning (HRL), a high-level policy specifies a subgoal for the low-level policy to reach. Effective HRL hinges on a suitable subgoal representation function, abstracting state space into latent subgoal space and inducing varied low-level behaviors. Exi…

2023

State-Conditioned Adversarial Subgoal Generation

AAAI 2023technical

Hierarchical reinforcement learning (HRL) proposes to solve difficult tasks by performing decision-making and control at successively higher levels of temporal abstraction. However, off-policy HRL often suffers from the problem of a non-stationary high-level policy since the low-level policy is cons…

Cited by 8SourcePDFScholar