← Search

Anoop Saladi

5 accepted papers

2025

AutoEval-ToD: Automated Evaluation of Task-oriented Dialog Systems

NAACL 2025long

Task-oriented Dialog systems (ToD) are essential in automating user interactions, but their complex design and dynamic nature make evaluation particularly challenging. Current evaluation methodologies heavily depend on human annotators, which can be inefficient, subjective, and expensive to scale. T…

Cited by 0SourcePDFScholar
2025

AutoKB: Automated Creation of Structured Knowledge Bases for Domain-Specific Support

NAACL 2025industry

Effective customer support requires domain-specific solutions tailored to users’ issues. However, LLMs like ChatGPT, while excelling in open-domain tasks, often face challenges such as hallucinations, lack of domain compliance, and imprecise solutions when applied to specialized contexts. RAG-based…

Cited by 0SourcePDFScholar
2025

VADE: Visual Attention Guided Hallucination Detection and Elimination

ACL 2025finding

Vision Language Models (VLMs) have achieved significant advancements in complex visual understanding tasks. However, VLMs are prone to hallucinations—generating outputs that lack alignment with visual content. This paper addresses hallucination detection in VLMs by leveraging the visual grounding in…

Cited by 0SourcePDFScholar
2025

VIT-Pro: Visual Instruction Tuning for Product Images

NAACL 2025industry

General vision-language models (VLMs) trained on web data struggle to understand and converse about real-world e-commerce product images. We propose a cost-efficient approach for collecting training data to train a generative VLM for e-commerce product images. The key idea is to leverage large-scale…

Cited by 0SourcePDFScholar
2024

Leveraging Uncertainty Estimates To Improve Classifier Performance

ICLR 2024poster

Binary classification typically involves predicting the label of an instance based on whether the model score for the positive class exceeds a threshold chosen based on the application requirements (e.g., maximizing recall for a precision bound). However, model scores are often not aligned with true…

Cited by 1SourcePDFScholar