← Search

Misha Sra

9 accepted papers

2026

CoDA: Agentic Systems for Collaborative Data Visualization

ICLR 2026poster

Automating data visualization from natural language is crucial for data science, yet current systems struggle with complex datasets containing multiple files and iterative refinement. Existing approaches, including simple single- or multi-agent systems, often oversimplify the task, focusing on initi…

Cited by 0SourcecodeScholar
2025

GraphEval36K: Benchmarking Coding and Reasoning Capabilities of Large Language Models on Graph Datasets

NAACL 2025findings

Large language models (LLMs) have achieved remarkable success in natural language processing (NLP), demonstrating significant capabilities in processing and understanding text data. However, recent studies have identified limitations in LLMs’ ability to manipulate, program, and reason about structur…

Cited by 0SourcePDFScholar
2025

Instruct-CLIP: Improving Instruction-Guided Image Editing with Automated Data Refinement Using Contrastive Learning

CVPR 2025poster

Although natural language instructions offer an intuitive way to guide automated image editing, deep-learning models often struggle to achieve high-quality results, largely due to the difficulty of creating large, high-quality training datasets. To do this, previous approaches have typically relied…

2025

TR-LLM: Integrating Trajectory Data for Scene-Aware LLM-Based Human Action Prediction

IROS 2025

Accurate prediction of human behavior is crucial for AI systems to effectively support real-world applications, such as autonomous robots anticipating and assisting with human tasks. Real-world scenarios frequently present challenges such as occlusions and incomplete scene observations, which can co

Cited by 4SourcecodeScholar
2024

AID-AppEAL: Automatic Image Dataset and Algorithm for Content Appeal Enhancement and Assessment Labeling

ECCV 2024poster

"We propose Image Content Appeal Assessment (), a novel metric that quantifies the level of positive interest an image’s content generates for viewers, such as the appeal of food in a photograph. This is fundamentally different from traditional Image-Aesthetics Assessment (IAA), which judges an imag…

2024

TiNO-Edit: Timestep and Noise Optimization for Robust Diffusion-Based Image Editing

CVPR 2024poster

Despite many attempts to leverage pre-trained text-to-image models (T2I) like Stable Diffusion (SD) for controllable image editing producing good predictable results remains a challenge. Previous approaches have focused on either fine-tuning pre-trained T2I models on specific datasets to generate ce…

2024

XplainLLM: A Knowledge-Augmented Dataset for Reliable Grounded Explanations in LLMs

EMNLP 2024main

Large Language Models (LLMs) have achieved remarkable success in natural language tasks, yet understanding their reasoning processes remains a significant challenge. We address this by introducing XplainLLM, a dataset accompanying an explanation framework designed to enhance LLM transparency and rel…

2022

Self-Supervised Knowledge Assimilation for Expert-Layman Text Style Transfer

AAAI 2022technical

Expert-layman text style transfer technologies have the potential to improve communication between members of scientific communities and the general public. High-quality information produced by experts is often filled with difficult jargon laypeople struggle to understand. This is a particularly not…