← Search

Xiaoxiao Guo

11 accepted papers

2026

GPR-GSLAM: Gaussian Process Regression-Enhanced Real-Time RGB-D SLAM Using Gaussian Splatting

RA-L 2026

3D Gaussian Splatting (3DGS) has recently revolutionized novel view synthesis and provided a new paradigm for photorealistic Simultaneous Localization and Mapping (SLAM). However, current 3DGS-based RGB-D SLAM systems still face three key challenges: incomplete depth observations due to sensor noise

Cited by 0SourceScholar
2023

JECC: Commonsense Reasoning Tasks Derived from Interactive Fictions

ACL 2023findings

Commonsense reasoning simulates the human ability to make presumptions about our physical world, and it is an essential cornerstone in building general AI systems. We proposea new commonsense reasoning dataset based on human’s Interactive Fiction (IF) gameplaywalkthroughs as human players demonstrat…

2021

Augmenting Policy Learning with Routines Discovered from a Single Demonstration

AAAI 2021technical

Humans can abstract prior knowledge from very little data and use it to boost skill learning. In this paper, we propose routine-augmented policy learning (RAPL), which discovers routines composed of primitive actions from a single demonstration and uses discovered routines to augment policy learning…

2021

Fashion IQ: A New Dataset Towards Retrieving Images by Natural Language Feedback

CVPR 2021poster

Conversational interfaces for the detail-oriented retail fashion domain are more natural, expressive, and user friendly than classical keyword-based search interfaces. In this paper, we introduce the Fashion IQ dataset to support and advance research on interactive fashion image retrieval. Fashion I…

Cited by 297PDFcodeScholar
2019

Drill-down: Interactive Retrieval of Complex Scenes using Natural Language Queries

NeurIPS 2019poster

This paper explores the task of interactive image retrieval using natural language queries, where a user progressively provides input queries to refine a set of retrieval results. Moreover, our work explores this problem in the context of complex image scenes containing multiple objects. We propose…

2018

Dialog-based Interactive Image Retrieval

NeurIPS 2018poster

Existing methods for interactive image retrieval have demonstrated the merit of integrating user feedback, improving retrieval results. However, most current systems rely on restricted forms of user feedback, such as binary relevance responses, or feedback based on a fixed set of relative attributes…

2018

Eigenoption Discovery through the Deep Successor Representation

ICLR 2018poster

Options in reinforcement learning allow agents to hierarchically decompose a task into subtasks, having the potential to speed up learning and planning. However, autonomously learning effective sets of options is still a major challenge in the field. In this paper we focus on the recently introduced…

Cited by 194SourcePDFScholar
2018

Evidence Aggregation for Answer Re-Ranking in Open-Domain Question Answering

ICLR 2018poster

Very recently, it comes to be a popular approach for answering open-domain questions by first searching question-related passages, then applying reading comprehension models to extract answers. Existing works usually extract answers from single passages independently, thus not fully make use of the…

2017

Learning to Query, Reason, and Answer Questions On Ambiguous Texts

ICLR 2017poster

A key goal of research in conversational systems is to train an interactive agent to help a user with a task. Human conversation, however, is notoriously incomplete, ambiguous, and full of extraneous detail. To operate effectively, the agent must not only understand what was explicitly conveyed but…

Cited by 31SourceScholar
2015

Action-Conditional Video Prediction using Deep Networks in Atari Games

NeurIPS 2015spotlight

Motivated by vision-based reinforcement learning (RL) problems, in particular Atari games from the recent benchmark Aracade Learning Environment (ALE), we consider spatio-temporal prediction problems where future (image-)frames are dependent on control variables or actions as well as previous frames…

Cited by 1061SourcePDFScholar