2025
FESTA: Functionally Equivalent Sampling for Trust Assessment of Multimodal LLMs
EMNLP 2025
The accurate trust assessment of multimodal large language models (MLLMs) generated predictions, which can enable selective prediction and improve user confidence, is challenging due to the diverse multi-modal input paradigms. We propose F unctionally E quivalent S ampling for T rust A ssessment (FE