← Search

Kento Sasaki

2 accepted papers

2026

STRIDE-QA: Visual Question Answering Dataset for Spatiotemporal Reasoning in Urban Driving Scenes

AAAI 2026technical

Vision-Language Models (VLMs) have been applied to autonomous driving to support decision-making in complex real-world scenarios. However, their training on static, web-sourced image-text pairs fundamentally limits the precise spatiotemporal reasoning required to understand and predict dynamic traff

Cited by 0SourcePDFScholar
2026

Understanding Sensitivity of Differential Attention through the Lens of Adversarial Robustness

ICLR 2026poster

Differential Attention (DA) has been proposed as a refinement to standard attention, suppressing redundant or noisy context through a subtractive structure and thereby reducing contextual hallucination. While this design sharpens task-relevant focus, we show that it also introduces a structural frag…

Cited by 0SourceScholar