← Search

Ahmed Hendawy

2 accepted papers

2026

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning

ICLR 2026poster

The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target network remains a compromise solution that preserves stability at the cost of slowly moving targets, thus delaying learning. Conversely, using the online n…

Cited by 0SourcecodeScholar
2024

Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts

ICLR 2024poster

Multi-Task Reinforcement Learning (MTRL) tackles the long-standing problem of endowing agents with skills that generalize across a variety of problems. To this end, sharing representations plays a fundamental role in capturing both unique and common characteristics of the tasks. Tasks may exhibit si…