← Search

Filipp Fisin

2 accepted papers

2026

LK Losses: Direct Acceptance Rate Optimization for Speculative Decoding

ICML 2026poster

Speculative decoding accelerates autoregressive large language model (LLM) inference by using a lightweight draft model to propose candidate tokens that are then verified in parallel by the target model. The speedup is significantly determined by the acceptance rate, yet standard training minimizes …

Cited by 0SourceScholar
2025

Guided Search Strategies in Non-Serializable Environments with Applications to Software Engineering Agents

ICML 2025poster

Large language models (LLMs) have recently achieved remarkable results in complex multi-step tasks, such as mathematical reasoning and agentic software engineering. However, they often struggle to maintain consistent performance across multiple solution attempts. One effective approach to narrow the…

Cited by 0SourcePDFScholar