← Search

Amit Sheth

17 accepted papers

2026

CausalPulse: Agentic Copilot for Root Cause Analysis in Smart Manufacturing

AAAI 2026technical

Modern manufacturing systems demand real-time, trustworthy, and interpretable insights into anomalies and their underlying causes. However, conventional pipelines treat anomaly detection, causal inference, and decision-making as siloed tasks, lacking integration, explainability, and adaptability. We

Cited by 0SourcePDFScholar
2026

CausalTrace: A Neurosymbolic Causal Analysis Agent for Smart Manufacturing

AAAI 2026technical

Modern manufacturing environments demand not only accurate predictions but also interpretable insights to process anomalies, root causes, and potential interventions. Existing AI systems often function as isolated black boxes, lacking the seamless integration of prediction, explanation, and causal r

Cited by 0SourcePDFScholar
2026

Chatsparent: An Interactive System for Detecting and Mitigating Cognitive Fatigue in LLMs

AAAI 2026technical

LLMs are increasingly being deployed as chatbots, but today’s interfaces offer little to no friction: users interact through seamless conversations that conceal when the model is drifting, hallucinating or failing. This lack of transparency fosters blind trust, even as models produce unstable or rep

Cited by 0SourcePDFScholar
2026

Cognitive Fatigue in Autoregressive Transformers: Formalization and Measurement

ICML 2026poster

Autoregressive language models frequently degrade during long-horizon generation, producing repetitive text, losing instruction adherence, and exhibiting unstable entropy. Despite the prevalence of these failures, practitioners lack online diagnostics to detect them in real time as they occur. We fo…

Cited by 0SourceScholar
2026

DETONATE – A Benchmark for Text-to-Image Alignment and Kernelized Direct Preference Optimization

AAAI 2026technical

Alignment is crucial for text-to-image (T2I) models to ensure that the generated images faithfully capture user intent while maintaining safety and fairness. Direct Preference Optimization (DPO) has emerged as a key alignment technique for large language models (LLMs), and its influence is now exten

Cited by 0SourcePDFScholar
2026

In-Situ Eval: A Modular Framework for Custom and Real-Time RAG Benchmarking

AAAI 2026technical

Retrieval-Augmented Generation (RAG) has become the standard approach for integrating domain knowledge into Large Language Models (LLMs). However, fair comparison of RAG pipelines remains difficult: data preparation is often ad hoc, subsampling methods are opaque, parameters vary across implementati

Cited by 0SourcePDFScholar
2026

PAL: Personal Adaptive Learner

AAAI 2026technical

AI-driven education platforms have made some progress in personalisation, yet most remain constrained to static adaptation—predefined quizzes, uniform pacing, or generic feedback—limiting their ability to respond to learners’ evolving understanding. This shortfall highlights the need for systems tha

Cited by 0SourcePDFScholar
2025

KnowledgePrompts: Exploring the Abilities of Large Language Models to Solve Proportional Analogies via Knowledge-Enhanced Prompting

COLING 2025main

Making analogies is fundamental to cognition. Proportional analogies, which consist of four terms, are often used to assess linguistic and cognitive abilities. For instance, completing analogies like “Oxygen is to Gas as < blank > is to < blank >" requires identifying the semantic relationship (e.g.…

2025

NSF-MAP: Neurosymbolic Multimodal Fusion for Robust and Interpretable Anomaly Prediction in Assembly Pipelines

IJCAI 2025

In modern assembly pipelines, identifying anomalies is crucial in ensuring product quality and operational efficiency. Conventional single-modality methods fail to capture the intricate relationships required for precise anomaly prediction in complex predictive environments with abundant data and mu

2025

YinYang-Align: A new Benchmark for Competing Objectives and Introducing Multi-Objective Preference based Text-to-Image Alignment

ACL 2025finding

Precise alignment in Text-to-Image (T2I) systems is crucial for generating visuals that reflect user intent while adhering to ethical and policy standards. Recent controversies, such as the Google Gemini-generated Pope image backlash, highlight the urgent need for robust alignment mechanisms. Buildi…

Cited by 0SourcePDFScholar
2024

GEAR-Up: Generative AI and External Knowledge-Based Retrieval: Upgrading Scholarly Article Searches for Systematic Reviews

AAAI 2024technical

This paper addresses the time-intensive nature of systematic reviews (SRs) and proposes a solution leveraging advancements in Generative AI (e.g., ChatGPT) and external knowledge augmentation (e.g., Retrieval-Augmented Generation). The proposed system, GEAR-Up, automates query development and transl…

Cited by 7SourcePDFScholar
2023

ANALOGICAL - A Novel Benchmark for Long Text Analogy Evaluation in Large Language Models

ACL 2023findings

Over the past decade, analogies, in the form of word-level analogies, have played a significant role as an intrinsic measure of evaluating the quality of word embedding methods such as word2vec. Modern large language models (LLMs), however, are primarily evaluated on extrinsic measures based on benc…

Cited by 34SourcePDFScholar
2023

CLUE-AD: A Context-Based Method for Labeling Unobserved Entities in Autonomous Driving Data

AAAI 2023technical

Generating high-quality annotations for object detection and recognition is a challenging and important task, especially in relation to safety-critical applications such as autonomous driving (AD). Due to the difficulty of perception in challenging situations such as occlusion, degraded weather, and…

Cited by 7SourcePDFScholar
2023

Demo Alleviate: Demonstrating Artificial Intelligence Enabled Virtual Assistance for Telehealth: The Mental Health Case

AAAI 2023technical

After the pandemic, artificial intelligence (AI) powered support for mental health care has become increasingly important. The breadth and complexity of significant challenges required to provide adequate care involve: (a) Personalized patient understanding, (b) Safety-constrained and medically vali…

Cited by 21SourcePDFScholar
2023

FACTIFY-5WQA: 5W Aspect-based Fact Verification through Question Answering

ACL 2023long

Automatic fact verification has received significant attention recently. Contemporary automatic fact-checking systems focus on estimating truthfulness using numerical scores which are not human-interpretable. A human fact-checker generally follows several logical steps to verify a verisimilitude cla…

2020

Identifying Depressive Symptoms from Tweets: Figurative Language Enabled Multitask Learning Framework

COLING 2020main

Existing studies on using social media for deriving mental health status of users focus on the depression detection task. However, for case management and referral to psychiatrists, health-care workers require practical and scalable depressive disorder screening and triage system. This study aims to…

Cited by 51SourcePDFScholar