← Search

Siqi Bao

8 accepted papers

2026

Enhancing Agentic Search via Data Synthesis on Hierarchical Constraint Satisfaction

ICLR 2026poster

Deep research becomes increasingly important as people seek to solve complex problems that require gathering and synthesizing information from diverse sources. A key capability in this process is agentic search, where an LLM-agent iteratively retrieves relevant information across multiple sources wh…

Cited by 0SourceScholar
2026

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

ICML 2026spotlight

The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bottleneck stemming from the complexity of constructing executable environments across diverse languages. To address this, we introduce **MEnvAgent**, a **M**ult…

Cited by 0SourceScholar
2026

ProxyAttn: Guided Sparse Attention via Representative Heads

ICLR 2026poster

The quadratic complexity of attention mechanisms limits the efficiency of Large Language Models (LLMs) on long-text tasks. Recently, methods that dynamically estimate block importance have enabled efficient block sparse attention, leading to significant acceleration in long-text pre-filling of LLMs.…

Cited by 0SourcecodeScholar
2026

Retro*: Optimizing LLMs for Reasoning-Intensive Document Retrieval

ICLR 2026poster

With the growing popularity of LLM agents and RAG, it has become increasingly important to retrieve documents that are essential for solving a task, even when their connection to the task is indirect or implicit. Addressing this problem requires fine-grained reasoning to accurately assess the releva…

Cited by 0SourceScholar
2023

Query Enhanced Knowledge-Intensive Conversation via Unsupervised Joint Modeling

ACL 2023long

In this paper, we propose an unsupervised query enhanced approach for knowledge-intensive conversations, namely QKConv. There are three modules in QKConv: a query generator, an off-the-shelf knowledge selector, and a response generator. QKConv is optimized through joint training, which produces the…

2023

Towards Boosting the Open-Domain Chatbot with Human Feedback

ACL 2023long

Many open-domain dialogue models pre-trained with social media comments can generate coherent replies but have difficulties producing engaging responses. This phenomenon might mainly result from the deficiency of annotated human-human conversations and the misalignment with human preference. In this…

2022

Q-TOD: A Query-driven Task-oriented Dialogue System

EMNLP 2022main

Existing pipelined task-oriented dialogue systems usually have difficulties adapting to unseen domains, whereas end-to-end systems are plagued by large-scale knowledge bases in practice. In this paper, we introduce a novel query-driven task-oriented dialogue system, namely Q-TOD. The essential infor…

2016

A unified framework for atlas-based segmentation with forward deformation and label refinement

ICASSP 2016accepted

In this paper, a novel unified framework for atlas-based segmentation is proposed, consisting of two main components: forward deformation and label refinement. A newly designed distance constraint on mesh edges is enforced with contrast sensitivity in forward deformation based on Markov random field…

Cited by 0SourceScholar