← Search

Yuhao Sun

17 accepted papers

2026

Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!

ICLR 2026poster

Fine-tuning on open-source Large Language Models (LLMs) with proprietary data is now a standard practice for downstream developers to obtain task-specific models. Surprisingly, we reveal a new and concerning risk along with the practice: the provider of the open-source LLMs can later extract the pri…

Cited by 0SourcecodeScholar
2026

Multiplicative Orthogonal Sequential Editing for Language Models

AAAI 2026technical

Knowledge editing aims to efficiently modify the internal knowledge of large language models (LLMs) without compromising their other capabilities. The prevailing editing paradigm, which appends an update matrix to the original parameter matrix, has been shown by some studies to damage key numerical

Cited by 0SourcePDFScholar
2026

Orthogonal Concept Erasure for Diffusion Models

ICML 2026oral

Concept erasure has emerged as a promising approach to mitigate undesired or unsafe content in diffusion models, yet existing methods still face significant limitations. While training-based methods are effective, their high computational cost limits scalability. Editing-based methods are more effic…

Cited by 0SourceScholar
2026

SDErasure: Concept-Specific Trajectory Shifting for Concept Erasure via Adaptive Diffusion Classifier

ICLR 2026poster

Concept erasure methods have proven effective in mitigating the potential for text‑to‑image diffusion models to produce harmful content. Nevertheless, prevailing methods based on post fine-tuning introduce substantial disruption to the original model’s parameter distribution and suffer from excessiv…

Cited by 0SourceScholar
2026

TangleScore: Tangle-Guided Purge and Imprint for Unstructured Knowledge Editing

ICLR 2026poster

Large language models (LLMs) struggle with inaccurate and outdated information, driving the emergence of knowledge editing as a lightweight alternative. Despite their effectiveness in modifying structured knowledge, existing editing methods often fail to generalize to unstructured cases, particularl…

Cited by 0SourceScholar
2026

WFR-FM: Simulation-Free Dynamic Unbalanced Optimal Transport

ICLR 2026poster

The Wasserstein–Fisher–Rao (WFR) metric extends dynamic optimal transport (OT) by coupling displacement with change of mass, providing a principled geometry for modeling unbalanced snapshot dynamics. Existing WFR solvers, however, are often unstable, computationally expensive, and difficult to scale…

Cited by 0SourcecodeScholar
2025

AnyTouch: Learning Unified Static-Dynamic Representation across Multiple Visuo-tactile Sensors

ICLR 2025poster

Visuo-tactile sensors aim to emulate human tactile perception, enabling robots to precisely understand and manipulate objects. Over time, numerous meticulously designed visuo-tactile sensors have been integrated into robotic systems, aiding in completing various tasks. However, the distinct data cha…

2025

Invisible Watermarks, Visible Gains: Steering Machine Unlearning with Bi-Level Watermarking Design

ICCV 2025poster

With the increasing demand for the right to be forgotten, machine unlearning (MU) has emerged as a vital tool for enhancing trust and regulatory compliance by enabling the removal of sensitive data influences from machine learning (ML) models. However, most MU algorithms primarily rely on in-trainin…

Cited by 0SourcePDFScholar
2025

Modeling Cell Dynamics and Interactions with Unbalanced Mean Field Schrödinger Bridge

NeurIPS 2025poster

Modeling the dynamics from sparsely time-resolved snapshot data is crucial for understanding complex cellular processes and behavior. Existing methods leverage optimal transport, Schrödinger bridge theory, or their variants to simultaneously infer stochastic, unbalanced dynamics from snapshot data.…

Cited by 0SourcecodeScholar
2025

SPO: Self Preference Optimization with Self Regularization

EMNLP 2025

Direct Preference Optimization (DPO) is a widely used offline preference optimization algorithm that enhances the simplicity and training stability of reinforcement learning through reward function reparameterization from PPO. Recently, SimPO (Simple Preference Optimization) and CPO (Contrastive Pre

Cited by 0SourcePDFScholar
2025

Sample-specific Noise Injection for Diffusion-based Adversarial Purification

ICML 2025poster

*Diffusion-based purification* (DBP) methods aim to remove adversarial noise from the input sample by first injecting Gaussian noise through a forward diffusion process, and then recovering the clean example through a reverse generative process. In the above process, how much Gaussian noise is injec…

2025

Variational Regularized Unbalanced Optimal Transport: Single Network, Least Action

NeurIPS 2025poster

Recovering the dynamics from a few snapshots of a high-dimensional system is a challenging task in statistical physics and machine learning, with important applications in computational biology. Many algorithms have been developed to tackle this problem, based on frameworks such as optimal transport…

Cited by 0SourcecodeScholar
2024

DiffAM: Diffusion-based Adversarial Makeup Transfer for Facial Privacy Protection

CVPR 2024poster

With the rapid development of face recognition (FR) systems the privacy of face images on social media is facing severe challenges due to the abuse of unauthorized FR systems. Some studies utilize adversarial attack techniques to defend against malicious FR systems by generating adversarial examples…

2023

Evolving Connectivity for Recurrent Spiking Neural Networks

NeurIPS 2023poster

Recurrent spiking neural networks (RSNNs) hold great potential for advancing artificial general intelligence, as they draw inspiration from the biological nervous system and show promise in modeling complex dynamics. However, the widely-used surrogate gradient-based training methods for RSNNs are in…

2023

Implementation and Optimization of Grasping Learning with Dual-modal Soft Gripper

ICRA 2023poster

Robust and efficient grasping of different objects is still an open problem due to the difficulty of integrating multidisciplinary knowledge such as gripper ontology design, perception, control, and learning. In recent years, learning-based methods have achieved excellent results in grasping various…

Cited by 4SourceScholar
2023

TIRgel: A Visuo-Tactile Sensor With Total Internal Reflection Mechanism for External Observation and Contact Detection

RA-L 2023

This letter proposes a vision-based tactile sensor named TIRgel, leveraging visual integration to simplify the sensing system. First, the sensor achieves conversion of visual and tactile modality via focus adjustment. Under far-focus imaging, the camera can observe external environments; Under near-

Cited by 36SourceScholar
2021

An Efficient Paper Anti-Counterfeiting Method Based on Microstructure Orientation Estimation

ICASSP 2021accepted

Forgery of paper invoices and certificates negatively causes huge economic loss every year. In the molding process of paper, plant fibres inside will distribute randomly and form the unique microscopic surface. Accordingly, forensic researchers have presented many techniques to build its feature des…

Cited by 0SourceScholar