← Search

Ruoxi Jiang

10 accepted papers

2026

DisPPO: Quantile-Based Distributional Reinforcement Learning for Large Language Models

ICML 2026poster

Reinforcement Learning (RL) has become a cornerstone for enhancing the reasoning capabilities of Large Language Models (LLMs). However, standard actor-critic methods, such as PPO, rely on scalar value functions that estimate only the expectation of cumulative returns. This reduction inherently disca…

Cited by 0SourceScholar
2026

PMDformer: Patch-Mean Decoupling Transformer for Long-term Forecasting

ICLR 2026poster

Long-term time series forecasting (LTSF) plays a crucial role in fields such as energy management, finance, and traffic prediction. Transformer-based models have adopted patch-based strategies to capture long-range dependencies, but accurately modeling shape similarities across patches and variables…

Cited by 0SourceScholar
2026

Residual Connections Harm Generative Representation Learning

CVPR 2026

We show that introducing a weighting factor to reduce the influence of identity shortcuts in residual networks significantly enhances semantic feature learning in generative representation learning frameworks, such as masked autoencoders (MAEs) and diffusion models. Our modification improves linear

Cited by 11SourcecodeScholar
2025

Hierarchical Implicit Neural Emulators

NeurIPS 2025poster

Neural PDE solvers offer a powerful tool for modeling complex dynamical systems, but often struggle with error accumulation over long time horizons and maintaining stability and physical consistency. We introduce a multiscale implicit neural emulator that enhances long-term prediction accuracy by co…

Cited by 0SourceScholar
2023

Training neural operators to preserve invariant measures of chaotic attractors

NeurIPS 2023poster

Chaotic systems make long-horizon forecasts difficult because small perturbations in initial conditions cause trajectories to diverge at an exponential rate. In this setting, neural operators trained to minimize squared error losses, while capable of accurate short-term forecasts, often fail to repr…

2022

Embed and Emulate: Learning to estimate parameters of dynamical systems with uncertainty quantification

NeurIPS 2022accept

This paper explores learning emulators for parameter estimation with uncertainty estimation of high-dimensional dynamical systems. We assume access to a computationally complex simulator that inputs a candidate parameter and outputs a corresponding multi-channel time series. Our task is to accuratel…

2021

Pure Exploration in Kernel and Neural Bandits

NeurIPS 2021poster

We study pure exploration in bandits, where the dimension of the feature representation can be much larger than the number of arms. To overcome the curse of dimensionality, we propose to adaptively embed the feature representation of each arm into a lower-dimensional space and carefully deal with th…

Cited by 22SourcePDFScholar