← Search

Xiaoqin Wang

7 accepted papers

2026

FineXtrol: Controllable Motion Generation via Fine-Grained Text

AAAI 2026technical

Recent works have sought to enhance the controllability and precision of text-driven motion generation. Some approaches leverage large language models (LLMs) to produce more detailed texts, while others incorporate global 3D coordinate sequences as additional control signals. However, the former oft

Cited by 0SourcePDFScholar
2026

Real-World Unsupervised Models Generalize to Predict Brain Responses to Out-of-Distribution Stimuli

ICML 2026spotlight

Deep neural networks currently provide the leading quantitative models of neural responses in sensory systems. However, these networks remain implausible as models of sensory development, largely because they rely on supervised training with label efficiency far exceeding that of biological learning…

Cited by 0SourceScholar
2025

EAGLE: Expert-Guided Self-Enhancement for Preference Alignment in Pathology Large Vision-Language Model

ACL 2025long

Recent advancements in Large Vision Language Models (LVLMs) show promise for pathological diagnosis, yet their application in clinical settings faces critical challenges of multimodal hallucination and biased responses. While preference alignment methods have proven effective in general domains, acq…

2025

FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs

CVPR 2025poster

Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in various tasks. However, effectively evaluating these MLLMs on face perception remains largely unexplored. To address this gap, we introduce FaceBench, a dataset featuring hierarchical multi-view and multi-level att…

2025

MentalGLM Series: Explainable Large Language Models for Mental Health Analysis on Chinese Social Media

EMNLP 2025

With the rise of mental health challenges, social media has become a key platform for emotional expression. Deep learning offers a promising solution for analyzing mental health but lacks flexibility and interpretability. Large language models (LLMs) introduce greater adaptability and can explain th

2024

Neural Embeddings Rank: Aligning 3D latent dynamics with movements

NeurIPS 2024poster

Aligning neural dynamics with movements is a fundamental goal in neuroscience and brain-machine interfaces. However, there is still a lack of dimensionality reduction methods that can effectively align low-dimensional latent dynamics with movements. To address this gap, we propose Neural Embeddings…

2015

Self-calibration in visual sensor networks equipped with RGB-D cameras

ICASSP 2015accepted

We consider the self-calibration problem (estimation of location and orientation of multiple camera sensors), in visual sensor networks equipped with RGB-D cameras. We propose two algorithms based on feature matching and relative pose estimation. First one uses Floyd-Warshall algorithm, and can accu…

Cited by 0SourceScholar