← Search

Kenji Doya

5 accepted papers

2025

Training Recurrent Neural Networks with Inherent Missing Data for Wearable Device Applications (Student Abstract)

AAAI 2025technical

Wearable devices are transforming healthcare by providing continuous, real-time physiological data for monitoring and analysis. However, data often suffer from noise and significant missing values due to operational constraints and user compliance. Traditional approaches address these issues through…

Cited by 0SourcePDFScholar
2022

Variational oracle guiding for reinforcement learning

ICLR 2022poster

How to make intelligent decisions is a central problem in machine learning and artificial intelligence. Despite recent successes of deep reinforcement learning (RL) in various decision making problems, an important but under-explored aspect is how to leverage oracle observation (the information that…

2019

Theoretical Analysis of Efficiency and Robustness of Softmax and Gap-Increasing Operators in Reinforcement Learning

AISTATS 2019poster

In this paper, we propose and analyze conservative value iteration, which unifies value iteration, soft value iteration, advantage learning, and dynamic policy programming. Our analysis shows that algorithms using a combination of gap-increasing and max operators are resilient to stochastic errors,…

Cited by 45SourcePDFScholar
2018

PIPPS: Flexible Model-Based Policy Search Robust to the Curse of Chaos

ICML 2018oral

Previously, the exploding gradient problem has been explained to be central in deep learning and model-based reinforcement learning, because it causes numerical issues and instability in optimization. Our experiments in model-based reinforcement learning imply that the problem is not just a numerica…