← Search

Yu-Chee Tseng

8 accepted papers

2026

Partial Ring Scan: Revisiting Scan Order in Vision State Space Models

ICML 2026poster

State Space Models (SSMs) have emerged as efficient alternatives to attention for vision tasks, offering linear-time sequence processing with competitive accuracy. Vision SSMs, however, require serializing 2D images into 1D token sequences along a predefined scan order, a factor often overlooked. We…

Cited by 0SourceScholar
2025

GCC: Generative Color Constancy via Diffusing a Color Checker

CVPR 2025poster

Color constancy methods often struggle to generalize across different camera sensors due to varying spectral sensitivities. We present GCC, which leverages diffusion models to inpaint color checkers into images for illumination estimation. Our key innovations include (1) a single-step deterministic…

Cited by 0SourcePDFScholar
2025

SDA-LLM: Spatial DisAmbiguation via Multi-turn Vision-Language Dialogues for Robot Navigation

IROS 2025

When users give natural language instructions to service robots, positional information is often referenced relative to objects in the environment rather than absolute coordinates. However, humans naturally use relative references. For example, in“Go to the chair and pick up empty bottles”, where th

Cited by 0SourceScholar
2025

SpectroMotion: Dynamic 3D Reconstruction of Specular Scenes

CVPR 2025poster

We present SpectroMotion, a novel approach that combines 3D Gaussian Splatting (3DGS) with physically-based rendering (PBR) and deformation fields to reconstruct dynamic specular scenes. Previous methods extending 3DGS to model dynamic scenes have struggled to represent specular surfaces accurately.…

2023

UPLIFT: Unsupervised Person Labeling and Identification via Cooperative Learning with Mobile Robots

ICRA 2023poster

As robots are widely used in assisting manual tasks, an interesting challenge is: Can mobile robots help create a labeled knowledge dataset that can be used for efficiently creating deep learning models for other sensors? This paper proposes an Unsupervised Person Labeling and Identification (UPLIFT…

Cited by 0SourceScholar
2019

Enabling Identity-Aware Tracking via Fusion of Visual and Inertial Features

ICRA 2019poster

Person identification and tracking (PIT) is an essential issue in computer vision and robotic applications. It has long been studied and achieved by technologies such as RFID or face/fingerprint/iris recognition. These approaches, however, have their limitations due to environmental constraints (suc…

Cited by 22SourceScholar
2019

Who Takes What: Using RGB-D Camera and Inertial Sensor for Unmanned Monitor

ICRA 2019poster

Advanced Internet of Things (IoT) techniques have made human-environment interaction much easier. Existing solutions usually enable such interactions without knowing the identities of action performers. However, identifying users who are interacting with environments is a key to enable personalized…

Cited by 3SourceScholar
2018

Eye on You: Fusing Gesture Data from Depth Camera and Inertial Sensors for Person Identification

ICRA 2018poster

Person identification (PID) is a key issue in many IoT applications. It has long been studied and achieved by technologies such as RFID and face/fingerprint/iris recognition. These approaches, however, have their limitations due to environmental constraints (such as lighting and obstacles) or requir…

Cited by 18SourceScholar