← Search

I-Hsin Chung

4 accepted papers

2025

Attention Tracker: Detecting Prompt Injection Attacks in LLMs

NAACL 2025findings

Large Language Models (LLMs) have revolutionized various domains but remain vulnerable to prompt injection attacks, where malicious inputs manipulate the model into ignoring original instructions and executing designated action. In this paper, we investigate the underlying mechanisms of these attack…

2025

Cost-Efficient LLM Training with Lifetime-Aware Tensor Offloading via GPUDirect Storage

NeurIPS 2025poster

We present the design and implementation of a new lifetime-aware tensor offloading framework for GPU memory expansion using low-cost PCIe-based solid-state drives (SSDs). Our framework, TERAIO, is developed explicitly for large language model (LLM) training with multiple GPUs and multiple SSDs. Its…

Cited by 0SourceScholar
2025

STAR: Spectral Truncation and Rescale for Model Merging

NAACL 2025short

Model merging is an efficient way of obtaining a multi-task model from several pretrained models without further fine-tuning, and it has gained attention in various domains, including natural language processing (NLP). Despite the efficiency, a key challenge in model merging is the seemingly inevita…

2024

Overload: Latency Attacks on Object Detection for Edge Devices

CVPR 2024poster

Nowadays the deployment of deep learning-based applications is an essential task owing to the increasing demands on intelligent services. In this paper we investigate latency attacks on deep learning applications. Unlike common adversarial attacks for misclassification the goal of latency attacks is…

Cited by 15SourcePDFScholar