← Search

Hongwei Fan

8 accepted papers

2026

AT-VLA: Adaptive Tactile Injection for Enhanced Feedback Reaction in Vision-Language-Action Models

CVPR 2026

Vision-Language-Action (VLA) models have significantly advanced robotic agents capable of executing diverse tasks; however, they remain limited in contact-rich manipulation scenarios that require precise physical interactions. To address this limitation, recent studies have attempted to incorporate

Cited by 0SourceScholar
2026

CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model

AAAI 2026technical

Existing vision-and-language navigation models often deviate from the correct trajectory when executing instructions. However, these models lack effective error correction capability, hindering their recovery from errors. To address this challenge, we propose Self-correction Flywheel, a novel post-t

Cited by 25SourcePDFScholar
2026

FreeArtGS: Articulated Gaussian Splatting Under Free-moving Scenario

CVPR 2026

The increasing demand for augmented reality and robotics is driving the need for articulated object reconstruction with high scalability. However, existing settings for reconstructing from discrete articulation states or casual monocular videos require non-trivial axis alignment or suffer from insuf

Cited by 0SourcecodeScholar
2026

Real2Edit2Real: Generating Robotic Demonstrations via a 3D Control Interface

CVPR 2026

Recent progress in robot learning has been driven by large-scale datasets and powerful visuomotor policy architectures, yet policy robustness remains limited by the substantial cost of collecting diverse demonstrations, particularly for spatial generalization in manipulation tasks. To reduce repetit

Cited by 0SourceScholar
2025

BiAssemble: Learning Collaborative Affordance for Bimanual Geometric Assembly

ICML 2025poster

Shape assembly, the process of combining parts into a complete whole, is a crucial skill for robots with broad real-world applications. Among the various assembly tasks, geometric assembly—where broken parts are reassembled into their original form (e.g., reconstructing a shattered bowl)—is particul…

Cited by 0SourcePDFScholar
2025

Interacted Object Grounding in Spatio-Temporal Human-Object Interactions

AAAI 2025technical

Spatio-temporal Human-Object Interaction (ST-HOI) understanding aims at detecting HOIs from videos, which is crucial for activity understanding. However, existing whole-body-object interaction video benchmarks overlook the truth that open-world objects are diverse, that is, they usually provide limi…

2025

SimLauncher: Launching Sample-Efficient Real-World Robotic Reinforcement Learning via Simulation Pre-Training

IROS 2025

Autonomous learning of dexterous, long-horizon robotic skills has been a longstanding pursuit of embodied AI. Recent advances in robotic reinforcement learning (RL) have demonstrated remarkable performance and robustness in real-world visuomotor control tasks. However, applying RL in the real world

Cited by 3SourceScholar
2023

BioMassters: A Benchmark Dataset for Forest Biomass Estimation using Multi-modal Satellite Time-series

NeurIPS 2023poster

Above Ground Biomass is an important variable as forests play a crucial role in mitigating climate change as they act as an efficient, natural and cost-effective carbon sink. Traditional field and airborne LiDAR measurements have been proven to provide reliable estimations of forest biomass. Neverth…