← Search

Shaan Ul Haque

2 accepted papers

2026

Fine-Tuning Diffusion Models via Intermediate Distribution Shaping

ICLR 2026poster

Diffusion models are widely used for generative tasks across domains. While pre-trained diffusion models effectively capture the training data distribution, it is often desirable to shape these distributions using reward functions to align with downstream applications. Policy gradient methods, such…

Cited by 0SourceScholar
2025

Stochastic Approximation with Unbounded Markovian Noise: A General-Purpose Theorem

AISTATS 2025poster

Motivated by engineering applications such as resource allocation in networks and inventory systems, we consider average-reward Reinforcement Learning with unbounded state space and reward function. Recent work Murthy et al. 2024 studied this problem in the actor-critic framework and established fin…

Cited by 0SourceScholar