← Search

Jingmin Chen

4 accepted papers

2026

Move-Then-Operate: Behavioral Phasing for Human-Like Robotic Manipulation

ICML 2026poster

We present Move-Then-Operate, a Vision–language–action framework that explicitly decouples robotic manipulation into two distinct behavioral phases: coarse relocation (move) and contact-critical interaction (operate). Unlike monolithic policies that conflate these heterogeneous regimes, our architec…

Cited by 0SourceScholar
2026

Task-Related Token Compression in Multimodal Large Language Models from an Explainability Perspective

ICLR 2026poster

Existing Multimodal Large Language Models (MLLMs) process a large number of visual tokens, leading to significant computational costs and inefficiency. Instruction-related visual token compression demonstrates strong task relevance, which aligns well with MLLMs’ ultimate goal of instruction followin…

Cited by 0SourceScholar
2025

Stimulating Imagination: Towards General-purpose "Something Something Placement"

IROS 2025

General-purpose object placement is a fundamental capability of an intelligent generalist robot: being capable of rearranging objects following precise human instructions even in novel environments. This work is dedicated to achieving general-purpose object placement with "something something" instr

Cited by 0SourceScholar
2021

Exploiting Behavioral Consistence for Universal User Representation

AAAI 2021technical

User modeling is critical for developing personalized services in industry. A common way for user modeling is to learn user representations that can be distinguished by their interests or preferences. In this work, we focus on developing universal user representation model. The obtained universal re…