← Search

Haozhe Chen

13 accepted papers

2026

Topology-aware Feature Propagation for Unsupervised Non-rigid Point Cloud Correspondence

CVPR 2026

Unsupervised non-rigid point cloud correspondence aims to predict point-to-point correspondences without annotations. Existing methods leverage the spatial-relation-based feature propagation strategy that includes non-physical connections, which are sensitive to non-rigid deformation. To address thi

Cited by 0SourceScholar
2025

Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian Framework

NeurIPS 2025poster

Careful curation of data sources can significantly improve the performance of LLM pre-training, but predominant approaches rely heavily on intuition or costly trial-and-error, making them difficult to generalize across different data domains and downstream tasks. Although scaling laws can provide a…

Cited by 0SourcecodeScholar
2025

DiT4Edit: Diffusion Transformer for Image Editing

AAAI 2025technical

Despite recent advances in UNet-based image editing, methods for shape-aware object editing in high-resolution images are still lacking. Compared to UNet, Diffusion Transformers (DiT) demonstrate superior capabilities to effectively capture the long-range dependencies among patches, leading to highe…

2025

Graph-Embedded Structure-Aware Perceptual Hashing for Neural Network Protection and Piracy Detection

CVPR 2025poster

The advancement of AI technology has significantly influenced production activities, increasing the focus on copyright protection for AI models. Model perceptual hashing offers an efficient solution for retrieving the pirated models. Existing methods, such as handcrafted feature-based and dual-branc…

Cited by 0SourcePDFScholar
2025

Human-like Navigation in a World Built for Humans

CoRL 2025poster

When navigating in a man-made environment they haven’t visited before—like an office building—humans employ behaviors such as reading signs and asking others for directions. These behaviors help humans reach their destinations efficiently by reducing the need to search through large areas. Existing…

Cited by 0SourceScholar
2025

RoboVerse: A Unified Platform, Benchmark and Dataset for Scalable and Generalizable Robot Learning

RSS 2025poster

Data scaling and standardized evaluation benchmarks have driven remarkable advances in natural language processing and computer vision. However, in robotics, scaling up data and establishing evaluation protocols pose significant challenges. Directly collecting real-world data is inefficient and reso…

Cited by 0PDFScholar
2024

INViTE: INterpret and Control Vision-Language Models with Text Explanations

ICLR 2024poster

Large-scale pre-trained vision foundation models, such as CLIP, have become de facto backbones for various vision tasks. However, due to their black-box nature, understanding the underlying rules behind these models’ predictions and controlling model behaviors have remained open challenges. We prese…

2024

QGym: Scalable Simulation and Benchmarking of Queuing Network Controllers

NeurIPS 2024poster

Queuing network control allows allocation of scarce resources to manage congestion, a fundamental problem in manufacturing, communications, and healthcare. Compared to standard RL problems, queueing problems are distinguished by unique challenges: i) a system operating in continuous time, ii) high…

2024

SelfIE: Self-Interpretation of Large Language Model Embeddings

ICML 2024poster

How do large language models (LLMs) obtain their answers? The ability to explain and control an LLM’s reasoning process is key for reliability, transparency, and future model developments. We propose SelfIE (Self-Interpretation of Embeddings), a framework that enables LLMs to interpret their own emb…

2022

Speech Pattern Based Black-Box Model Watermarking for Automatic Speech Recognition

ICASSP 2022accepted

As an effective method for intellectual property (IP) protection, model watermarking technology has been applied on a wide variety of deep neural networks (DNN), including speech classification models. However, how to design a black-box watermarking scheme for automatic speech recognition (ASR) mode…

Cited by 0SourceScholar