← Search

Guangtao Lyu

6 accepted papers

2026

Beyond Global Alignment: Fine-Grained Motion-Language Retrieval via Pyramidal Shapley-Taylor Learning

ICML 2026spotlight

As a foundational task in human-centric cross-modal intelligence, motion-language retrieval aims to bridge the semantic gap between natural language and human motion, enabling intuitive motion analysis, yet existing approaches predominantly focus on aligning entire motion sequences with global textu…

Cited by 0SourceScholar
2026

Channel-masked Asymmetric Distribution Matching for Cross-Domain Generalized Dataset Distillation

AAAI 2026technical

Dataset distillation has achieved remarkable progress as an effective approach for data compression. However, real-world data often comes from diverse domains, leading to potential mismatches between the domains of synthesized images and those of the evaluation set. Existing methods primarily assume

Cited by 0SourcePDFScholar
2026

Learning Attribute–Affordance Hierarchies in Hyperbolic Space for Open-Vocabulary 3D Object Affordance Grounding

ICML 2026poster

This paper pays attention to open-vocabulary 3D object affordance grounding (OVAG), which aims to localize affordance regions on 3D objects by leveraging interaction images or textual instructions. Most existing methods treat interaction images as sources of external affordance knowledge and align t…

Cited by 0SourceScholar
2025

Smooth and Flexible Camera Movement Synthesis via Temporal Masked Generative Modeling

NeurIPS 2025poster

In dance performances, choreographers define the visual expression of movement, while cinematographers shape its final presentation through camera work. Consequently, the synthesis of camera movements informed by both music and dance has garnered increasing research interest. While recent advancemen…

Cited by 0SourceScholar
2025

Towards Unified Human Motion-Language Understanding via Sparse Interpretable Characterization

ICLR 2025poster

Recently, the comprehensive understanding of human motion has been a prominent area of research due to its critical importance in many fields. However, existing methods often prioritize specific downstream tasks and roughly align text and motion features within a CLIP-like framework. This results in…

Cited by 1SourcePDFScholar
2024

LLM Knows Body Language, Too: Translating Speech Voices into Human Gestures

ACL 2024long

In response to the escalating demand for digital human representations, progress has been made in the generation of realistic human gestures from given speeches. Despite the remarkable achievements of recent research, the generation process frequently includes unintended, meaningless, or non-realist…

Cited by 3SourcePDFScholar