← Search

Weiyi Lu

5 accepted papers

2026

SeGO: Sensitivity-Aware Golden Optimization for Large-Scale VLM Quantization

IJCAI 2026

The deployment of Vision-Language Models (VLMs) faces memory and computational bottlenecks because of the massive parameters and intensive computations. While Post-Training Quantization (PTQ) can reduce these costs, existing methods often overlook the heterogeneity of multimodal input when applied t

Cited by 0Scholar
2026

TranTac: Leveraging Transient Tactile Signals for Contact-Rich Robotic Manipulation

ICRA 2026poster

Robotic manipulation tasks such as inserting a key into a lock or plugging a USB device into a port can fail when visual perception is insufficient to detect misalignment. In these situations, touch sensing is crucial for the robot to monitor the task's states and make precise, timely adjustments. C…

2022

Asynchronous Convergence in Multi-Task Learning via Knowledge Distillation from Converged Tasks

NAACL 2022industry

Multi-task learning (MTL) aims to solve multiple tasks jointly by sharing a base representation among them. This can lead to more efficient learning and better generalization, as compared to learning each task individually. However, one issue that often arises in MTL is the convergence speed between…

Cited by 4SourcePDFScholar
2022

DynaMaR: Dynamic Prompt with Mask Token Representation

EMNLP 2022industry

Recent research has shown that large language models pretrained using unsupervised approaches can achieve significant performance improvement on many downstream tasks. Typically when adapting these language models to downstream tasks, like a classification or regression task, we employ a fine-tuning…

Cited by 1SourcePDFScholar
2021

Top-Down Attention in End-to-End Spoken Language Understanding

ICASSP 2021accepted

Spoken language understanding (SLU) is the task of inferring the semantics of spoken utterances. Traditionally, this has been achieved with a cascading combination of Automatic Speech Recognition (ASR) and Natural Language Understanding (NLU) modules that are optimized separately, which can lead to…

Cited by 0SourceScholar