← Search

Hengheng Zhang

4 accepted papers

2026

Learn from Your Mistakes: Tree-like Self-Play on Vulnerability Nodes for Secure Code LLMs

ICML 2026poster

While Large Language Models (LLMs) excel in code generation, they remain prone to replicating subtle yet critical vulnerabilities endemic to their training data. Current alignment techniques, such as Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL), typically apply coarse-grained optimiz…

Cited by 0SourceScholar
2024

QA-LoRA: Quantization-Aware Low-Rank Adaptation of Large Language Models

ICLR 2024poster

Recently years have witnessed a rapid development of large language models (LLMs). Despite the strong ability in many language-understanding tasks, the heavy computational burden largely restricts the application of LLMs especially when one needs to deploy them onto edge devices. In this paper, we p…

2020

Rethinking the Distribution Gap of Person Re-identification with Camera-based Batch Normalization

ECCV 2020poster

The fundamental difficulty in person re-identification (ReID) lies in learning the correspondence among individual cameras. It strongly demands costly inter-camera annotations, yet the trained models are not guaranteed to transfer well to previously unseen cameras. These problems significantly limit…

2017

Person Re-Identification in the Wild

CVPR 2017spotlight

This paper presents a novel large-scale dataset and comprehensive baselines for end-to-end pedestrian detection and person recognition in raw video frames. Our baselines address three issues: the performance of various combinations of detectors and recognizers, mechanisms for pedestrian detection to…

Cited by 1011PDFcodeScholar