← Search

Junhao Li

4 accepted papers

2026

Closing the Safety Gap: Surgical Concept Erasure in Visual Autoregressive Models

ICLR 2026poster

The rapid progress of visual autoregressive (VAR) models has brought new opportunities for text-to-image generation, but also heightened safety concerns. Existing concept erasure techniques, primarily designed for diffusion models, fail to generalize to VARs due to their next-scale token prediction…

Cited by 0SourcecodeScholar
2026

DARC: Dual Adjustment Reasoning with Counterfactuals for Trustworthy Chest X-ray Classification

CVPR 2026

Despite their impressive performance in multi-label classification of chest X-ray images (CXR), deep learning models are widely plagued by two types of spurious correlations: feature confounding arising from pathological co-occurrence and shortcut learning triggered by non-pathological visual confou

Cited by 0SourceScholar
2026

From ``Sure" to ``Sorry": Detecting Jailbreak in Large Vision Language Model via JailNeurons

ICLR 2026poster

Large Vision-Language Models (LVLMs) are vulnerable to jailbreak attacks that can generate harmful content. Existing detection methods are either limited to detecting specific attack types or are too time-consuming, making them impractical for real-world deployment. To address these challenges, we p…

Cited by 0SourcecodeScholar
2026

Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation

ICRA 2026poster

Current 3D-aware pretraining methods for embodied perception and manipulation are largely built on differentiable rendering frameworks, producing either fully implicit neural fields or fully explicit geometric primitives. Implicit representations, while expressive, lack explicit structural cues, whe…