← Search

Feng Han

11 accepted papers

2026

CG-THWM: Curriculum-Guided Temporal Haptic World Modeling for Peg-In-Hole Tasks

ICRA 2026poster

Fine-tolerance peg-in-hole manipulation demands high precision under contact-rich, nonsmooth dynamics, where irregular geometries, inclinations, and tight-clearance interference often cause model-free reinforcement learning (RL) to fail. We propose the Curriculum-Guided Temporal Haptic World Model (…

Cited by 0Scholar
2026

GrOCE : Graph-Guided Online Concept Erasure for Text-to-Image Diffusion Models

CVPR 2026

Concept erasure aims to remove harmful, inappropriate, or copyrighted content from text-to-image diffusion models while preserving non-target semantics. However, existing methods either rely on costly fine-tuning or apply coarse semantic separation, often degrading unrelated concepts and lacking ada

Cited by 0SourcecodeScholar
2026

Twin-DP3: View-Invariant 3D Diffusion Policy With Digital Twin

RA-L 2026

Learning visuomotor policies with imitation learning from 3D observations is a primary research direction in robotic manipulation, as 3D data inherently captures spatial features critical for action. While many existing methods rely on multiview point cloud fusion, recent studies like 3D Diffusion P

Cited by 0SourceScholar
2025

DuMo: Dual Encoder Modulation Network for Precise Concept Erasure

AAAI 2025technical

The exceptional generative capability of text-to-image models has raised substantial safety concerns regarding the generation of Not-Safe-For-Work (NSFW) content and potential copyright infringement. To address these concerns, previous methods safeguard the models by eliminating inappropriate concep…

2024

Gaussian Process-Enhanced, External and Internal Convertible Form-Based Control of Underactuated Balance Robots

ICRA 2024poster

External and internal convertible (EIC) form-based motion control (i.e., EIC-based control) is one of the effective approaches for underactuated balance robots. By sequentially controller design, trajectory tracking of the actuated subsystem and balance of the unactuated subsystem can be achieved si…

Cited by 1SourceScholar
2024

PLaD: Preference-based Large Language Model Distillation with Pseudo-Preference Pairs

ACL 2024findings

Large Language Models (LLMs) have exhibited impressive capabilities in various tasks, yet their vast parameter sizes restrict their applicability in resource-constrained settings. Knowledge distillation (KD) offers a viable solution by transferring expertise from large teacher models to compact stud…

Cited by 5SourcePDFScholar
2024

VIEWS: Entity-Aware News Video Captioning

EMNLP 2024main

Existing popular video captioning benchmarks and models often produce generic captions for videos that lack specific identification of individuals, locations, or organizations (named entities). However, in the case of news videos, the setting is more demanding, requiring the inclusion of such named…

2022

Scaling Multimodal Pre-Training via Cross-Modality Gradient Harmonization

NeurIPS 2022accept

Self-supervised pre-training recently demonstrates success on large-scale multimodal data, and state-of-the-art contrastive learning methods often enforce the feature consistency from cross-modality inputs, such as video/audio or video/text pairs. Despite its convenience to formulate and leverage in…

Cited by 13SourcePDFScholar