← Search

Guofeng Quan

3 accepted papers

2025

RASD: Retrieval-Augmented Speculative Decoding

ACL 2025finding

Speculative decoding accelerates inference in large language models (LLMs) by generating draft tokens for target model verification. Current approaches for obtaining draft tokens rely on lightweight draft models or additional model structures to generate draft tokens and retrieve context from databa…

Cited by 0SourcePDFScholar
2022

Is MultiWOZ a Solved Task? An Interactive TOD Evaluation Framework with User Simulator

EMNLP 2022finding

Task-Oriented Dialogue (TOD) systems are drawing more and more attention in recent studies.Current methods focus on constructing pre-trained models or fine-tuning strategies while the evaluation of TOD is limited by a policy mismatch problem.That is, during evaluation, the user utterances are from t…