← Search

Zhiwei Ge

3 accepted papers

2025

Unraveling the Mystery: Defending Against Jailbreak Attacks Via Unearthing Real Intention

COLING 2025main

As Large Language Models (LLMs) become more advanced, the security risks they pose also increase. Ensuring that LLM behavior aligns with human values, particularly in mitigating jailbreak attacks with elusive and implicit intentions, has become a significant challenge. To address this issue, we prop…

2024

Exploring Region-Word Alignment in Built-in Detector for Open-Vocabulary Object Detection

CVPR 2024poster

Open-vocabulary object detection aims to detect novel categories that are independent from the base categories used during training. Most modern methods adhere to the paradigm of learning vision-language space from a large-scale multi-modal corpus and subsequently transferring the acquired knowledge…

Cited by 6SourcePDFScholar