← Search

Hongqiao Chen

2 accepted papers

2026

Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models

ICLR 2026poster

Spatio-temporal reasoning is a remarkable capability of Vision Language Models (VLMs), but the underlying mechanisms of such abilities remain largely opaque. We postulate that visual/geometrical and textual representations of spatial structure must be combined at some point in VLM computations. We s…

Cited by 0SourcecodeScholar
2026

Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

ICML 2026poster

Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. Existing attacks typically rely on fixed textual triggers and single-entity backdoor targets, making them highly susceptible to enumeration-based input defense…

Cited by 0SourceScholar