← Search

Khoa Doan

8 accepted papers

2026

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

ICML 2026poster

Vision–Language–Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-shot imitation learning remains limited. We conduct a systematic stress test of state-of-the-art VLA models and show that performance degrades sharply …

Cited by 0SourceScholar
2026

Retrospective Feature Estimation for Continual Learning

ICML 2026poster

The intrinsic capability to continuously learn a changing data stream is a desideratum of deep neural networks (DNNs). However, current DNNs suffer from catastrophic forgetting, which interferes with remembering past knowledge. To mitigate this issue, existing Continual Learning (CL) approaches ofte…

Cited by 0SourcecodeScholar
2026

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching

ICML 2026poster

Direct Preference Optimization (DPO) is a widely used RL-free method for aligning language models from pairwise preferences, but it models preferences over full sequences even though generation is driven by per-token decisions. Existing token-level extensions typically decompose a sequence-level Bra…

Cited by 0SourceScholar
2024

Cold-start Recommendation by Personalized Embedding Region Elicitation

UAI 2024poster

Rating elicitation is a success element for recommender systems to perform well at cold-starting, in which the systems need to recommend items to a newly arrived user with no prior knowledge about the user’s preference. Existing elicitation methods employ a fixed set of items to learn the user’s pre…

Cited by 0SourcePDFScholar
2024

Fooling the Textual Fooler via Randomizing Latent Representations

ACL 2024findings

Despite outstanding performance in a variety of Natural Language Processing (NLP) tasks, recent studies have revealed that NLP models are vulnerable to adversarial attacks that slightly perturb the input to cause the models to misbehave. Several attacks can even compromise the model without requirin…

Cited by 0SourcePDFScholar