← Search

Anna Arias-Duart

2 accepted papers

2026

Language Models Can Explain Visual Features via Steering

CVPR 2026

Sparse Autoencoders uncover thousands of features in vision models, yet explaining these features without requiring human intervention remains an open challenge. While previous work has proposed generating correlation-based explanations based on top activating input examples, we present a fundamenta

Cited by 0SourcecodeScholar
2025

Automatic Evaluation of Healthcare LLMs Beyond Question-Answering

NAACL 2025short

Current Large Language Models (LLMs) benchmarks are often based on open-ended or close-ended QA evaluations, avoiding the requirement of human labor. Close-ended measurements evaluate the factuality of responses but lack expressiveness. Open-ended capture the model’s capacity to produce discourse re…

Cited by 1SourcePDFScholar