← Search

Henrik Metternich

1 accepted papers

2026

Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning

ICLR 2026poster

The use of target networks is a popular approach for estimating value functions in deep Reinforcement Learning (RL). While effective, the target network remains a compromise solution that preserves stability at the cost of slowly moving targets, thus delaying learning. Conversely, using the online n…

Cited by 0SourcecodeScholar