← Search

Igor Margulis

1 accepted papers

2025

FastDraft: How to Train Your Draft

ACL 2025finding

Speculative Decoding has gained popularity as an effective technique for accelerating the auto-regressive inference process of Large Language Models. However, Speculative Decoding entirely relies on the availability of efficient draft models, which are often lacking for many existing language models…