← Search

Ravi Tej Akella

3 accepted papers

2024

Reasoning with Latent Diffusion in Offline Reinforcement Learning

ICLR 2024poster

Offline reinforcement learning (RL) holds promise as a means to learn high-reward policies from a static dataset, without the need for further environment interactions. However, a key challenge in offline RL lies in effectively stitching portions of suboptimal trajectories from the static dataset wh…

2021

Deep Bayesian Quadrature Policy Optimization

AAAI 2021technical

We study the problem of obtaining accurate policy gradient estimates using a finite number of samples. Monte-Carlo methods have been the default choice for policy gradient estimation, despite suffering from high variance in the gradient estimates. On the other hand, more sample efficient alternative…

2020

Reinforced Multi-task Approach for Multi-hop Question Generation

COLING 2020main

Question generation (QG) attempts to solve the inverse of question answering (QA) problem by generating a natural language question given a document and an answer. While sequence to sequence neural models surpass rule-based systems for QG, they are limited in their capacity to focus on more than one…

Cited by 24SourcePDFScholar