← Search

Yihua Cheng

11 accepted papers

2026

FoSS: Modeling Long-Range Dependencies and Multimodal Uncertainty in Trajectory Prediction via Fourier-State Space Integration

CVPR 2026

Accurate trajectory prediction is vital for safe autonomous driving, yet existing approaches struggle to balance modeling power and computational efficiency. Attention-based architectures incur quadratic complexity with increasing agents, while recurrent models struggle to capture long-range depende

Cited by 0SourceScholar
2026

Force-Aware 3D Contact Modeling for Stable Grasp Generation

AAAI 2026technical

Contact-based grasp generation plays a crucial role in various applications. Recent methods typically focus on the geometric structure of objects, producing grasps with diverse hand poses and plausible contact points. However, these approaches often overlook the physical attributes of the grasp, spe

Cited by 0SourcePDFScholar
2026

RTGaze: Real-Time 3D-Aware Gaze Redirection from a Single Image

AAAI 2026technical

Gaze redirection methods aim to generate realistic human face images with controllable eye movement. However, recent methods often struggle with 3D consistency, efficiency, or quality, limiting their practical applications. In this work, we propose RTGaze, a real-time and high-quality gaze redirecti

Cited by 0SourcePDFScholar
2025

3D Prior Is All You Need: Cross-Task Few-shot 2D Gaze Estimation

CVPR 2025poster

3D and 2D gaze estimation share the fundamental objective of capturing eye movements but are traditionally treated as two distinct research domains. In this paper, we introduce a novel cross-task few-shot 2D gaze estimation approach, aiming to adapt a pre-trained 3D gaze estimation network for 2D ga…

Cited by 0SourcePDFScholar
2025

PersonaBooth: Personalized Text-to-Motion Generation

CVPR 2025poster

This paper introduces Motion Personalization, a new task that generates personalized motions aligned with text descriptions using several basic motions containing Persona. To support this novel task, we introduce a new large-scale motion dataset called PerMo (PersonaMotion), which captures the uniqu…

Cited by 1SourcePDFScholar
2025

Single-view Image to Novel-view Generation for Hand-Object Interactions

AAAI 2025technical

Hand-object interaction modeling from a single RGB image is a significantly challenging task. Previous works typically reconstruct hand-object interactions as texture-less meshes, ignoring photo-realistic image generation. In this work, we introduce the HO123, a novel method to synthesize novel-view…

Cited by 0SourcePDFScholar
2025

Trajectory Mamba: Efficient Attention-Mamba Forecasting Model Based on Selective SSM

CVPR 2025poster

Motion prediction is crucial for autonomous driving, as it enables accurate forecasting of future vehicle trajectories based on historical inputs. This paper introduces Trajectory Mamba, a novel efficient trajectory prediction framework based on the selective state-space model (SSM). Conventional at…

2024

What Do You See in Vehicle? Comprehensive Vision Solution for In-Vehicle Gaze Estimation

CVPR 2024poster

Driver's eye gaze holds a wealth of cognitive and intentional cues crucial for intelligent vehicles. Despite its significance research on in-vehicle gaze estimation remains limited due to the scarcity of comprehensive and well-annotated datasets in real driving scenarios. In this paper we present th…

Cited by 10SourcePDFScholar
2023

DVGaze: Dual-View Gaze Estimation

ICCV 2023poster

Gaze estimation methods estimate gaze from facial appearance with a single camera. However, due to the limited view of a single camera, the captured facial appearance cannot provide complete facial information and thus complicate the gaze estimation problem. Recently, camera devices are rapidly upda…

Cited by 26PDFcodeScholar