← Search

Quang Nguyen

8 accepted papers

2026

Accelerated Sinkhorn Algorithms for Partial Optimal Transport

ICASSP 2026poster

Partial Optimal Transport (POT) addresses the problem of transporting only a fraction of the total mass between two distributions, making it suitable when marginals have unequal size or contain outliers. While Sinkhorn-based methods are widely used, their complexity bounds for POT remain suboptimal…

Cited by 0SourcePDFScholar
2026

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation

ICML 2026poster

We present \textbf{ExCyTIn-Bench}, the first benchmark to \textbf{E}valuate an LLM agent \textbf{X} on the task of \textbf{Cy}ber \textbf{T}hreat \textbf{In}vestigation through security questions derived from investigation graphs. Real‑world security analysts must sift through a large number of hete…

Cited by 0SourceScholar
2026

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

ICML 2026poster

Vision–Language–Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-shot imitation learning remains limited. We conduct a systematic stress test of state-of-the-art VLA models and show that performance degrades sharply …

Cited by 0SourceScholar
2026

SIGMA: A Physics-Based Benchmark for Gas Chimney Understanding in Seismic Images

CVPR 2026

Seismic images reconstruct subsurface reflectivity from field recordings, guiding exploration and reservoir monitoring. Gas chimneys are vertical anomalies caused by subsurface fluid migration. Understanding these phenomena is crucial for assessing hydrocarbon potential and avoiding drilling hazards

Cited by 0SourcecodeScholar
2025

CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models

ICCV 2025poster

Disentangling content and style from a single image, known as content-style decomposition (CSD), enables recontextualization of extracted content and stylization of extracted styles, offering greater creative flexibility in visual synthesis. While recent personalization methods have explored the dec…

Cited by 0SourcePDFScholar
2025

EgoMusic-driven Human Dance Motion Estimation with Skeleton Mamba

ICCV 2025poster

Estimating human dance motion is a challenging task with various industrial applications. Recently, many efforts have focused on predicting human dance motion using either egocentric video or music as input. However, the task of jointly estimating human motion from both egocentric video and music re…

Cited by 0SourcePDFScholar
2025

GraspMAS: Zero-Shot Language-driven Grasp Detection with Multi-Agent System

IROS 2025

Language-driven grasp detection has the potential to revolutionize human-robot interaction by allowing robots to understand and execute grasping tasks based on natural language commands. However, existing approaches face two key challenges. First, they often struggle to interpret complex text instru

Cited by 0SourcecodeScholar
2025

SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion

CVPR 2025poster

Recent advances in text-guided image editing enable users to perform image edits through simple text inputs, leveraging the extensive priors of multi-step diffusion-based text-to-image models. However, these methods often fall short of the speed demands required for real-world and on-device applicat…