← Search

Vatsal Malaviya

2 accepted papers

2025

AcT2I: Evaluating and Improving Action Depiction in Text-to-Image Models

EMNLP 2025

Text-to-Image (T2I) models have recently achieved remarkable success in generating images from textual descriptions. However, challenges still persist in accurately rendering complex scenes where actions and interactions form the primary semantic focus. Our key observation in this work is that T2I m

2025

Mars-Bench: A Benchmark for Evaluating Foundation Models for Mars Science Tasks

NeurIPS 2025poster

Foundation models have enabled rapid progress across many specialized domains by leveraging large-scale pre-training on unlabeled data, demonstrating strong generalization to a variety of downstream tasks. While such models have gained significant attention in fields like Earth Observation, their ap…

Cited by 0SourcecodeScholar