← Search

Fuchang Liu

1 accepted papers

2026

Paper Folding Puzzles: Can Multimodal Large Language Models Perform Spatial Reasoning?

AAAI 2026technical

Multimodal Large Language Models (MLLMs) largely lag human-level performance on abstract visual reasoning (AVR), which requires models to infer latent rules from visual question sets and generalize them to novel scenarios. Most AVR benchmarks are constrained to narrow and repetitive 2D patterns, inv

Cited by 0SourcePDFScholar