← Search

Gabriele Serussi

3 accepted papers

2026

HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering

CVPR 2026

Video Large Language Models (Video-LLMs) are improving rapidly, yet current Video Question Answering (VideoQA) benchmarks often admit single-cue shortcuts, under-testing reasoning that must integrate evidence across time. We introduce HERBench, a benchmark designed to make multi-evidence integration

Cited by 0SourcecodeScholar
2026

Structured Diffusion Bridges: Inductive Bias for Denoising Diffusion Bridges

ICML 2026poster

Modality translation is inherently under-constrained, as multiple cross-modal mappings may yield the same marginals. Recent work has shown that diffusion bridges are effective for this task. However, most existing approaches rely on fully paired datasets, thereby imposing a single data-driven constr…

Cited by 0SourceScholar
2024

Active propulsion noise shaping for multi-rotor aircraft localization

IROS 2024poster

Multi-rotor aerial autonomous vehicles (MAVs) primarily rely on vision for navigation purposes. However, visual localization and odometry techniques suffer from poor performance in low or direct sunlight, a limited field of view, and vulnerability to occlusions. Acoustic sensing can serve as a compl…

Cited by 0SourcecodeScholar