← Search

Kaijun Zhou

2 accepted papers

2026

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

ICML 2026poster

Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under tight cost and energy budgets. Most prior evaluations rely on desktop-grade GPUs, obscuring the trade-offs and opportunities offered by heterogeneous e…

Cited by 0SourceScholar
2025

ASDSV: Multimodal Generation Made Efficient with Approximate Speculative Diffusion and Speculative Verification

NeurIPS 2025poster

Diffusion in transformer is central to advances in high-quality multimodal generation but suffer from high inference latency due to their iterative nature. Inspired by speculative decoding's success in accelerating large language models, we propose Approximate Speculative Diffusion with Speculati…

Cited by 0SourceScholar