← Search

Yerram Varun

3 accepted papers

2025

Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts?

ICLR 2025poster

Large Language Models (LLMs) are known to be susceptible to crafted adversarial attacks or jailbreaks that lead to the generation of objectionable content despite being aligned to human preferences using safety fine-tuning methods. While the large dimensionality of input token space makes it inevita…

Cited by 2SourcePDFScholar
2024

FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts

ACL 2024findings

Existing benchmarks for visual question answering lack in visual grounding and complexity, particularly in evaluating spatial reasoning skills. We introduce FlowVQA, a novel benchmark aimed at assessing the capabilities of visual question-answering multimodal language models in reasoning with flowch…

Cited by 11SourcePDFScholar
2024

Time-Reversal Provides Unsupervised Feedback to LLMs

NeurIPS 2024spotlight

Large Language Models (LLMs) are typically trained to predict in the forward direction of time. However, recent works have shown that prompting these models to look back and critique their own generations can produce useful feedback. Motivated by this, we explore the question of whether LLMs can be…

Cited by 0SourcePDFScholar