← Search

Ping He

6 accepted papers

2026

Better Datasets Start from RefineLab: Automatic Optimization for High-Quality Dataset Refinement

AAAI 2026technical

High‑quality Question–Answer (QA) datasets are foundational for reliable Large Language Model (LLM) evaluation, yet even expert‑crafted datasets exhibit persistent gaps in domain coverage, misaligned difficulty distributions, and factual inconsistencies. The recent surge in generative model-powered

Cited by 0SourcePDFScholar
2026

HogVul: Black-box Adversarial Code Generation Framework Against LM-based Vulnerability Detectors

AAAI 2026technical

Recent advances in software vulnerability detection have been driven by Language Model (LM)-based approaches. However, these models remain vulnerable to adversarial attacks that exploit lexical and syntax perturbations, allowing critical flaws to evade detection. Existing black-box attacks on LM-bas

Cited by 0SourcePDFScholar
2026

S3Net: Spatiotemporally Separated Sparse Network for Neuromorphic Vision Processing

AAAI 2026technical

Dynamic Vision Sensor (DVS) asynchronously records sparse events triggered by changes in pixel intensity, offering high temporal resolution and low latency. Existing frame-based methods process event data densely, violating its inherent sparsity and introducing computational redundancy. While asynch

Cited by 0SourcePDFScholar
2025

CLMTracing: Black-box User-level Watermarking for Code Language Model Tracing

EMNLP 2025

With the widespread adoption of open-source code language models (code LMs), intellectual property (IP) protection has become an increasingly critical concern. While current watermarking techniques have the potential to identify the code LM to protect its IP, they have limitations when facing the mo

Cited by 0SourcePDFScholar
2024

BaDExpert: Extracting Backdoor Functionality for Accurate Backdoor Input Detection

ICLR 2024poster

We present a novel defense, against backdoor attacks on Deep Neural Networks (DNNs), wherein adversaries covertly implant malicious behaviors (backdoors) into DNNs. Our defense falls within the category of post-development defenses that operate independently of how the model was generated. The propo…

2024

Logarithmic Lenses: Exploring Log RGB Data for Image Classification

CVPR 2024poster

The design of deep network architectures and training methods in computer vision has been well-explored. However in almost all cases the images have been used as provided with little exploration of pre-processing steps beyond normalization and data augmentation. Virtually all images posted on the we…

Cited by 7SourcePDFScholar