← Search

Yuting Ye

11 accepted papers

2026

Beyond Vision: A Multimodal Dataset and Framework for Pest Recognition via Plant Electrophysiological Signals

IJCAI 2026

Precise pest identification is essential for sustainable agriculture. Current visual recognition systems are brittle in the wild, where performance degrades due to occlusion and variable illumination. In contrast, plant electrophysiological signals serve as a robust, all-weather physiological modali

Cited by 0Scholar
2025

EgoLM: Multi-Modal Language Model of Egocentric Motions

CVPR 2025poster

As wearable devices become more prevalent, understanding the user's motion is crucial for improving contextual AI systems. We introduce EgoLM, a versatile framework designed for egocentric motion understanding using multi-modal data. EgoLM integrates the rich contextual information from egocentric v…

Cited by 4SourcePDFScholar
2025

From Sparse Signal to Smooth Motion: Real-Time Motion Generation with Rolling Prediction Models

CVPR 2025poster

In extended reality (XR), generating full-body motion of the users is important to understand their actions, drive their virtual avatars for social interaction, and convey a realistic sense of presence. While prior works focused on spatially sparse and always-on input signals from motion controllers…

Cited by 0SourcePDFScholar
2024

Nymeria: A Massive Collection of Egocentric Multi-modal Human Motion in the Wild

ECCV 2024poster

"We introduce - a large-scale, diverse, richly annotated human motion dataset collected in the wild with multiple multimodal egocentric devices. The dataset comes with a) full-body ground-truth motion; b) multiple multimodal egocentric data from Project Aria devices with videos, eye tracking, IMUs a…

2023

PhaseMP: Robust 3D Pose Estimation via Phase-conditioned Human Motion Prior

ICCV 2023poster

We present a novel motion prior, called PhaseMP, modeling a probability distribution on pose transitions conditioned by a frequency domain feature extracted from a periodic autoencoder. The phase feature further enforces the pose transitions to be unidirectional (i.e. no backward movement in time),…

Cited by 21PDFScholar
2022

From Intervention to Domain Transportation: A Novel Perspective to Optimize Recommendation

ICLR 2022poster

The interventional nature of recommendation has attracted increasing attention in recent years. It particularly motivates researchers to formulate learning and evaluating recommendation as causal inference and data missing-not-at-random problems. However, few take seriously the consequence of violat…

Cited by 5SourcePDFScholar
2020

Fully Convolutional Mesh Autoencoder using Efficient Spatially Varying Kernels

NeurIPS 2020poster

Learning latent representations of registered meshes is useful for many 3D tasks. Techniques have recently shifted to neural mesh autoencoders. Although they demonstrate higher precision than traditional methods, they remain unable to capture fine-grained deformations. Furthermore, these methods can…

Cited by 97SourcePDFScholar
2018

Learning Warped Guidance for Blind Face Restoration

ECCV 2018poster

This paper studies the problem of blind face restoration from an unconstrained blurry, noisy, low-resolution, or compressed image (i.e., degraded observation). For better recovery of fine facial details, we modify the problem setting by taking both the degraded observation and a high-quality guided…