← Search

Huiyu Zhou

13 accepted papers

2026

CROWn: A Unified Framework for Anti-Aliased Downsampling and Phase-Calibrated Fusion in 3D Medical Segmentation

CVPR 2026

Precise 3D medical image segmentation is a clinical cornerstone for diagnosis, therapy planning, and longitudinal monitoring. However, routine acquisition with anisotropic voxel spacing and heterogeneous reconstruction induces downsampling aliasing and cross-scale misalignment that blur boundaries,

Cited by 0SourcecodeScholar
2026

Physically-Guided Optical Inversion Enable Non-Contact Side-Channel Attack on Isolated Screens

ICLR 2026poster

Noncontact exfiltration of electronic screen content poses a security challenge, with side-channel incursions as the principal vector. We introduce an optical projection side-channel paradigm that confronts two core instabilities: (i) the near-singular Jacobian spectrum of projection mapping breache…

Cited by 0SourceScholar
2026

Wavefront-Constrained Passive Obscured Object Detection

AAAI 2026technical

Accurately localizing and segmenting obscured objects from faint light patterns beyond the field of view is highly challenging due to multiple scattering and medium-induced perturbations. Most existing methods, based on real-valued modeling or local convolutional operations, are inadequate for captu

Cited by 0SourcePDFScholar
2025

EssayJudge: A Multi-Granular Benchmark for Assessing Automated Essay Scoring Capabilities of Multimodal Large Language Models

ACL 2025finding

Automated Essay Scoring (AES) plays a crucial role in educational assessment by providing scalable and consistent evaluations of writing tasks. However, traditional AES systems face three major challenges: (i) reliance on handcrafted features that limit generalizability, (ii) difficulty in capturing…

Cited by 0SourcePDFScholar
2025

RealRAG: Retrieval-augmented Realistic Image Generation via Self-reflective Contrastive Learning

ICML 2025poster

Recent text-to-image generative models, e.g., Stable Diffusion V3 and Flux, have achieved notable progress. However, these models are strongly restricted to their limited knowledge, a.k.a., their own fixed parameters, that are trained with closed datasets. This leads to significant hallucinations or…

Cited by 4SourcePDFScholar
2025

Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models

ACL 2025finding

Hallucination issues continue to affect multimodal large language models (MLLMs), with existing research mainly addressing object-level or attribute-level hallucinations, neglecting the more complex relation hallucinations that require advanced reasoning. Current benchmarks for relation hallucinatio…

2025

Segment Anyword: Mask Prompt Inversion for Open-Set Grounded Segmentation

ICML 2025poster

Open-set image segmentation poses a significant challenge because existing methods often demand extensive training or fine-tuning and generally struggle to segment unified objects consistently across diverse text reference expressions. Motivated by this, we propose Segment Anyword, a novel training-…

Cited by 0SourcePDFScholar
2025

Unlocking Speech Instruction Data Potential with Query Rewriting

ACL 2025finding

End-to-end Large Speech Language Models (**LSLMs**) demonstrate strong potential in response latency and speech comprehension capabilities, showcasing general intelligence across speech understanding tasks. However, the ability to follow speech instructions has not been fully realized due to the lac…

Cited by 0SourcePDFScholar
2024

D4-VTON: Dynamic Semantics Disentangling for Differential Diffusion based Virtual Try-On

ECCV 2024poster

"In this paper, we introduce D4 -VTON, an innovative solution for image-based virtual try-on. We address challenges from previous studies, such as semantic inconsistencies before and after garment warping, and reliance on static, annotation-driven clothing parsers. Additionally, we tackle the comple…

2024

Online Mouse Behavior Detection by Historical Dependency and Typical Instances

ICASSP 2024accepted

Mouse behavior analysis plays a pivotal role in the research of numerous neurodegenerative diseases. In this paper, we develop a novel online mouse behavior detection approach, which can recognize mice behaviors in real-time videos and pinpoint the initiation and cessation points of target behaviors…

Cited by 0SourceScholar
2022

Absolute Wrong Makes Better: Boosting Weakly Supervised Object Detection via Negative Deterministic Information

IJCAI 2022poster

Weakly supervised object detection (WSOD) is a challenging task, in which image-level labels (e.g., categories of the instances in the whole image) are used to train an object detector. Many existing methods follow the standard multiple instance learning (MIL) paradigm and have achieved promising pe…

Cited by 16SourcePDFScholar
2022

Semantically Contrastive Learning for Low-Light Image Enhancement

AAAI 2022technical

Low-light image enhancement (LLE) remains challenging due to the unfavorable prevailing low-contrast and weak-visibility problems of single RGB images. In this paper, we respond to the intriguing learning-related question -- if leveraging both accessible unpaired over/underexposed images and high-le…

2019

Score-specific Non-maximum Suppression and Coexistence Prior for Multi-scale Face Detection

ICASSP 2019accepted

Face detection is an ultimate component to support various visual facial related tasks. However, detecting faces with extremely low resolution or high occlusion is still an open problem. In this paper, we propose a two-step general approach to refine the performance of modern face detectors accordin…

Cited by 0SourceScholar