← Search

Guoping Zhao

3 accepted papers

2024

LI4: Label-Infused Iterative Information Interacting Based Fact Verification in Question-answering Dialogue

COLING 2024main

Fact verification constitutes a pivotal application in the effort to combat the dissemination of disinformation, a concern that has recently garnered considerable attention. However, previous studies in the field of fact verification, particularly those focused on question-answering dialogue, have e…

2024

Time-Efficient Reinforcement Learning with Stochastic Stateful Policies

ICLR 2024poster

Stateful policies play an important role in reinforcement learning, such as handling partially observable environments, enhancing robustness, or imposing an inductive bias directly into the policy structure. The conventional method for training stateful policies is Backpropagation Through Time (BPTT…

Cited by 3SourcePDFScholar
2023

LS-IQ: Implicit Reward Regularization for Inverse Reinforcement Learning

ICLR 2023poster

Recent methods for imitation learning directly learn a $Q$-function using an implicit reward formulation rather than an explicit reward function. However, these methods generally require implicit reward regularization to improve stability and often mistreat absorbing states. Previous works show that…