← Search

Alessandro Bondielli

2 accepted papers

2025

All-in-one: Understanding and Generation in Multimodal Reasoning with the MAIA Benchmark

EMNLP 2025

We introduce MAIA (Multimodal AI Assessment), a native-Italian benchmark designed for fine-grained investigation of the reasoning abilities of visual language models on videos. MAIA differs from other available video benchmarks for its design, its reasoning categories, the metric it uses, and the la

2025

ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models

ACL 2025finding

Large Language Models (LLMs) are increasingly used in tasks requiring interpretive and inferential accuracy. In this paper, we introduce ExpliCa, a new dataset for evaluating LLMs in explicit causal reasoning. ExpliCa uniquely integrates both causal and temporal relations presented in different ling…