← Search

Brian Park

2 accepted papers

2025

A Drop-In Solution for On-the-Fly Adaptation of Speculative Decoding in Large Language Models

ACL 2025long

Large Language Models (LLMs) are cutting-edge generative AI models built on transformer architecture, which tend to be highly memory-intensive when performing real-time inference. Various strategies have been developed to enhance the end-to-end inference speed for LLMs, one of which is speculative d…

Cited by 0SourcePDFScholar
2023

mBEST: Realtime Deformable Linear Object Detection Through Minimal Bending Energy Skeleton Pixel Traversals

RA-L 2023

Robotic manipulation of deformable materials is a challenging task that often requires realtime visual feedback. This is especially true for deformable linear objects (DLOs) or “rods”, whose slender and flexible structures make proper tracking and detection nontrivial. To address this challenge, we

Cited by 28SourcecodeScholar