← Search

Xingyu Wang

11 accepted papers

2026

CUBic: Coordinated Unified Bimanual Perception and Control Framework

CVPR 2026

Recent advances in visuomotor policy learning have enabled robots to perform control directly from visual inputs. Yet, extending such end-to-end learning from single-arm to bimanual manipulation remains challenging due to the need for both independent perception and coordinated interaction between a

Cited by 0SourceScholar
2026

UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routing and Test-Time Scaling

ICML 2026poster

In real-world deployments of large language models (LLMs), balancing inference quality and computational cost has become a central challenge. Existing approaches tackle this trade-off along two largely independent dimensions: model routing, which switches among models of different scales to match re…

Cited by 0SourceScholar
2025

Chain-of-Talkers (CoTalk): Fast Human Annotation of Dense Image Captions

EMNLP 2025

While densely annotated image captions significantly facilitate the learning of robust vision-language alignment, methodologies for systematically optimizing human annotation efforts remain underexplored. We introduce Chain-of-Talkers (CoTalk), an AI-in-the-loop methodology designed to maximize the

Cited by 0SourcePDFScholar
2025

DocMMIR: A Framework for Document Multi-modal Information Retrieval

EMNLP 2025

The rapid advancement of unsupervised representation learning and large-scale pre-trained vision-language models has significantly improved cross-modal retrieval tasks. However, existing multi-modal information retrieval (MMIR) studies lack a comprehensive exploration of document-level retrieval and

Cited by 0SourcePDFScholar
2025

Logic-of-Thought: Injecting Logic into Contexts for Full Reasoning in Large Language Models

NAACL 2025long

Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks but their performance in complex logical reasoning tasks remains unsatisfactory. Although some prompting methods, such as Chain-of-Thought, can improve the reasoning ability of LLMs to some extent, they suffe…

Cited by 9SourcePDFScholar
2025

RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation

RSS 2025poster

Developing robust and general-purpose manipulation policies is a key goal in robotics. To achieve effective generalization, it is essential to construct comprehensive datasets that encompass a large number of demonstration trajectories and diverse tasks. Unlike vision or language data, which can be…

Cited by 20PDFScholar
2023

Drift doesn't Matter: Dynamic Decomposition with Diffusion Reconstruction for Unstable Multivariate Time Series Anomaly Detection

NeurIPS 2023poster

Many unsupervised methods have recently been proposed for multivariate time series anomaly detection. However, existing works mainly focus on stable data yet often omit the drift generated from non-stationary environments, which may lead to numerous false alarms. We propose **D**ynamic **D**ecomposi…

2016

Removal of EEG artifacts for BCI applications using fully Bayesian tensor completion

ICASSP 2016accepted

High accuracy of electroencephalogram (EEG) classification can hardly be achieved if the signals are contaminated by severe artefacts. One helpless way to avoid such artefacts is usually to directly discard the severely disturbed EEG segments. This study considers a more elegant way that tries to re…

Cited by 0SourceScholar