← Search

Zhicong Huang

8 accepted papers

2026

MOAI: Module-Optimizing Architecture for Non-Interactive Secure Transformer Inference

ICLR 2026poster

Privacy concerns have been raised in Large Language Models (LLM) inference when models are deployed in Cloud Service Providers (CSP). Homomorphic encryption (HE) offers a promising solution by enabling secure inference directly over encrypted inputs. However, the high computational overhead of HE re…

Cited by 0SourcecodeScholar
2026

REVIS: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models

ICML 2026poster

Despite the advanced capabilities of Large Vision-Language Models (LVLMs), they frequently suffer from object hallucination. One reason is that visual features and pretrained textual representations often become intertwined in the deeper network layers. To address this, we propose REVIS, a training-…

Cited by 0SourceScholar
2026

Semantic Router: On the Feasibility of Hijacking MLLMs via a Single Adversarial Perturbation

ICML 2026poster

Multimodal Large Language Models (MLLMs) are increasingly deployed in stateless systems, such as autonomous driving and robotics. This paper investigates a novel threat: Semantic-Aware Hijacking. We explore the feasibility of hijacking multiple stateless decisions simultaneously using a single unive…

Cited by 0SourceScholar
2026

TEAR: Temporal-aware Automated Red-teaming for Text-to-Video Models

CVPR 2026

Text-to-Video (T2V) models are capable of synthesizing high-quality, temporally coherent dynamic video content, but the diverse generation also inherently introduces critical safety challenges. Existing safety evaluation methods, which focus on static image and text generation, are insufficient to c

Cited by 0SourceScholar
2025

ARUNet: Advancing Real-Time Stereo Matching for Robotic Perception on Edge Devices

RA-L 2025

Accurate real-time 3D depth perception is crucial for robotic systems, enabling navigation, obstacle avoidance, and object manipulation. However, the high computational demands of stereo matching networks hinder their application on resource-constrained robotic platforms. This paper presents ARUNet,

Cited by 2SourceScholar
2025

MARS: A Malignity-Aware Backdoor Defense in Federated Learning

NeurIPS 2025poster

Federated Learning (FL) is a distributed paradigm aimed at protecting participant data privacy by exchanging model parameters to achieve high-quality model training. However, this distributed nature also makes FL highly vulnerable to backdoor attacks. Notably, the recently proposed state-of-the-art…

Cited by 0SourceScholar
2021

Fast Light-Field Disparity Estimation With Multi-Disparity-Scale Cost Aggregation

ICCV 2021poster

Light field images contain both angular and spatial information of captured light rays. The rich information of light fields enables straightforward disparity recovery capability but demands high computational cost as well. In this paper, we design a lightweight disparity estimation model with physi…

Cited by 51PDFcodeScholar