← Search

Yufei Han

12 accepted papers

2025

Dissecting Logical Reasoning in LLMs: A Fine-Grained Evaluation and Supervision Study

EMNLP 2025

Logical reasoning is a core capability for large language models (LLMs), yet existing benchmarks that rely solely on final-answer accuracy fail to capture the quality of the reasoning process. To address this, we introduce FineLogic, a fine-grained evaluation framework that assesses logical reasonin

2025

PolGS: Polarimetric Gaussian Splatting for Fast Reflective Surface Reconstruction

ICCV 2025poster

Efficient shape reconstruction for surfaces with complex reflectance properties is crucial for real-time virtual reality. While 3D Gaussian Splatting (3DGS)-based methods offer fast novel view rendering by leveraging their explicit surface representation, their reconstruction quality lags behind tha…

Cited by 0SourcePDFScholar
2024

Attack-free Evaluating and Enhancing Adversarial Robustness on Categorical Data

ICML 2024poster

Research on adversarial robustness has predominantly focused on continuous inputs, leaving categorical inputs, especially tabular attributes, less examined. To echo this challenge, our work aims to evaluate and enhance the robustness of classification over categorical attributes against adversarial…

2024

BadRL: Sparse Targeted Backdoor Attack against Reinforcement Learning

AAAI 2024technical

Backdoor attacks in reinforcement learning (RL) have previously employed intense attack strategies to ensure attack success. However, these methods suffer from high attack costs and increased detectability. In this work, we propose a novel approach, BadRL, which focuses on conducting highly sparse b…

2024

Defending Jailbreak Prompts via In-Context Adversarial Game

EMNLP 2024main

Large Language Models (LLMs) demonstrate remarkable capabilities across diverse applications. However, concerns regarding their security, particularly the vulnerability to jailbreak attacks, persist. Drawing inspiration from adversarial training in deep learning and LLM agent learning processes, we…

2024

NeRSP: Neural 3D Reconstruction for Reflective Objects with Sparse Polarized Images

CVPR 2024poster

We present NeRSP a Neural 3D reconstruction technique for Reflective surfaces with Sparse Polarized images. Reflective surface reconstruction is extremely challenging as specular reflections are view-dependent and thus violate the multiview consistency for multiview stereo. On the other hand sparse…

Cited by 7SourcePDFScholar
2023

Poisoning with Cerberus: Stealthy and Colluded Backdoor Attack against Federated Learning

AAAI 2023technical

Are Federated Learning (FL) systems free from backdoor poisoning with the arsenal of various defense strategies deployed? This is an intriguing problem with significant practical implications regarding the utility of FL services. Despite the recent flourish of poisoning-resilient FL methods, our stu…

2023

Towards Efficient and Domain-Agnostic Evasion Attack with High-Dimensional Categorical Inputs

AAAI 2023technical

Our work targets at searching feasible adversarial perturbation to attack a classifier with high-dimensional categorical inputs in a domain-agnostic setting. This is intrinsically a NP-hard knapsack problem where the exploration space becomes explosively larger as the feature dimension increases. W…

2022

Towards Understanding the Robustness Against Evasion Attack on Categorical Data

ICLR 2022poster

Characterizing and assessing the adversarial vulnerability of classification models with categorical input has been a practically important, while rarely explored research problem. Our work echoes the challenge by first unveiling the impact factors of adversarial vulnerability of classification mode…

Cited by 10SourcePDFScholar