← Search

Shiqi Liu

12 accepted papers

2026

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

ICML 2026poster

Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond passive instruction following and making them essential for real-world, user-centric applications. Agentic reinforcement learning (RL) has recently eme…

Cited by 0SourceScholar
2025

Dynamics as Prompts: In-Context Learning for Sim-to-Real System Identifications

RA-L 2025

Sim-to-real transfer remains a significant challenge in robotics due to the discrepancies between simulated and real-world dynamics. Traditional methods like Domain Randomization often fail to capture fine-grained dynamics, limiting their effectiveness for precise control tasks. In this work, we pro

Cited by 15SourceScholar
2025

Learning Multi-Agent Loco-Manipulation for Long-Horizon Quadrupedal Pushing

ICRA 2025

Recently, quadrupedal locomotion has achieved significant success, but their manipulation capabilities, particularly in handling large objects, remain limited, restricting their usefulness in demanding real-world applications such as search and rescue, construction, industrial automation, and room o

Cited by 24SourceScholar
2025

LocoTouch: Learning Dynamic Quadrupedal Transport with Tactile Sensing

CoRL 2025poster

Quadrupedal robots have demonstrated remarkable agility and robustness in traversing complex terrains. However, they struggle with dynamic object interactions, where contact must be precisely sensed and controlled. To bridge this gap, we present LocoTouch, a system that equips quadrupedal robots wit…

Cited by 0SourceScholar
2025

One Filters All: A Generalist Filter For State Estimation

NeurIPS 2025poster

Estimating hidden states in dynamical systems, also known as optimal filtering, is a long-standing problem in various fields of science and engineering. In this paper, we introduce a general filtering framework, $\textbf{LLM-Filter}$, which leverages large language models (LLMs) for state estimation…

Cited by 0SourceScholar
2025

QuietPaw: Learning Quadrupedal Locomotion with Versatile Noise Preference Alignment

IROS 2025

When operating at their full capacity, quadrupedal robots can produce loud footstep noise, which can be disruptive in human-centered environments like homes, offices, and hospitals. As a result, balancing locomotion performance with noise constraints is crucial for the successful real-world deployme

Cited by 1SourceScholar
2025

Reverse Convolution and Its Applications to Image Restoration

ICCV 2025poster

Convolution and transposed convolution are fundamental operators widely used in neural networks. However, transposed convolution (a.k.a. deconvolution) does not serve as a true inverse of convolution due to inherent differences in their mathematical formulations. To date, no reverse convolution oper…

2024

Detecting AI-Generated Sentences in Human-AI Collaborative Hybrid Texts: Challenges, Strategies, and Insights

IJCAI 2024poster

This study explores the challenge of sentence-level AI-generated text detection within human-AI collaborative hybrid texts (abbreviated as hybrid texts). Existing studies of AI-generated text detection for hybrid texts often rely on synthetic datasets. These typically involve hybrid texts with a lim…

2024

OASIS: Conditional Distribution Shaping for Offline Safe Reinforcement Learning

NeurIPS 2024poster

Offline safe reinforcement learning (RL) aims to train a policy that satisfies con- straints using a pre-collected dataset. Most current methods struggle with the mismatch between imperfect demonstrations and the desired safe and rewarding performance. In this paper, we mitigate this issue from a da…

2023

Continual Vision-based Reinforcement Learning with Group Symmetries

CoRL 2023oral

Continual reinforcement learning aims to sequentially learn a variety of tasks, retaining the ability to perform previously encountered tasks while simultaneously developing new policies for novel tasks. However, current continual RL approaches overlook the fact that certain tasks are identical unde…

Cited by 10SourceScholar
2023

SeasonDepth: Cross-Season Monocular Depth Prediction Dataset and Benchmark Under Multiple Environments

IROS 2023poster

Different environments pose a great challenge to the outdoor robust visual perception for long-term autonomous driving, and the generalization of learning-based algorithms on different environments is still an open problem. Although monocular depth prediction has been well studied recently, few work…

Cited by 20SourcecodeScholar
2023

What Went Wrong? Closing the Sim-to-Real Gap via Differentiable Causal Discovery

CoRL 2023poster

Training control policies in simulation is more appealing than on real robots directly, as it allows for exploring diverse states in an efficient manner. Yet, robot simulators inevitably exhibit disparities from the real-world \rebut{dynamics}, yielding inaccuracies that manifest as the dynamical si…

Cited by 33SourceScholar