← Search

Sof{\'i}a Martinelli

2 accepted papers

2025

CaMMT: Benchmarking Culturally Aware Multimodal Machine Translation

EMNLP 2025

Translating cultural content poses challenges for machine translation systems due to the differences in conceptualizations between cultures, where language alone may fail to convey sufficient context to capture region-specific meanings. In this work, we investigate whether images can act as cultural

2025

HESEIA: A community-based dataset for evaluating social biases in large language models, co-designed in real school settings in Latin America

EMNLP 2025

Most resources for evaluating social biases in Large Language Models are developed without co-design from the communities affected by these biases, and rarely involve participatory approaches. We introduce HESEIA, a dataset of 46,499 sentences created in a professional development course. The course

Cited by 0SourcePDFScholar