← Search

Ali Saheb pasand

3 accepted papers

2026

Stable Deep Reinforcement Learning via Isotropic Gaussian Representations

ICML 2026spotlight

Deep reinforcement learning systems often suffer from unstable training dynamics due to non-stationarity, where learning objectives and data distributions evolve over time. We show that under non-stationary targets, isotropic Gaussian embeddings are provably advantageous. In particular, they induce …

Cited by 0SourceScholar
2022

Pro-KD: Progressive Distillation by Following the Footsteps of the Teacher

COLING 2022main

With the ever growing scale of neural models, knowledge distillation (KD) attracts more attention as a prominent tool for neural model compression. However, there are counter intuitive observations in the literature showing some challenging limitations of KD. A case in point is that the best perform…

Cited by 14SourcePDFScholar