← Search

Jaeyoon Kim

7 accepted papers

2026

Explaining Jailbreaks: Structured and Interpretable Safety Assessment for Large Language Models

IJCAI 2026

Large Language Models (LLMs) remain highly vulnerable to jailbreak attacks, yet existing evaluations rely primarily on outcome-level metrics such as Attack Success Rate (ASR), providing limited insight into how and why safety failures occur. We propose an explanation-aware safety framework that augm

Cited by 0Scholar
2026

GraphShield: Graph-Theoretic Modeling of Network-Level Dynamics for Robust Jailbreak Detection

ICLR 2026poster

Large language models (LLMs) are increasingly deployed in real-world applications but remain highly vulnerable to jailbreak prompts that bypass safety guardrails and elicit harmful outputs. We propose GraphShield, a graph-theoretic jailbreak detector that models information routing inside the LLM as…

Cited by 0SourceScholar
2026

Radiometrically Consistent Gaussian Surfels for Inverse Rendering

ICLR 2026oral

Inverse rendering with Gaussian Splatting has advanced rapidly, but accurately disentangling material properties from complex global illumination effects, particularly indirect illumination, remains a major challenge. Existing methods often query indirect radiance from Gaussian primitives pre-traine…

Cited by 1SourcecodeScholar
2026

Towards Test-time Efficient Visual Place Recognition via Asymmetric Query Processing

AAAI 2026technical

Visual Place Recognition (VPR) has advanced significantly with high-capacity foundation models like DINOv2, achieving remarkable performance. Nonetheless, their substantial computational cost makes deployment on resource-constrained devices impractical. In this paper, we introduce an efficient asymm

Cited by 0SourcePDFScholar
2024

Generalizable Person Re-identification via Balancing Alignment and Uniformity

NeurIPS 2024poster

Domain generalizable person re-identification (DG re-ID) aims to learn discriminative representations that are robust to distributional shifts. While data augmentation is a straightforward solution to improve generalization, certain augmentations exhibit a polarized effect in this task, enhancing in…