← Search

Ruchir Puri

4 accepted papers

2025

Agent Trajectory Explorer: Visualizing and Providing Feedback on Agent Trajectories

AAAI 2025technical

Agentic systems interleave large language model (LLM) reasoning, tool usage, and tool observations over multiple iterations to tackle complex tasks. The raw data from an agent's problem-solving process (the agents' trajectory) is not an ideal format for human analysis and oversight. There is a need…

Cited by 0SourcePDFScholar
2025

ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasks

ICML 2025oral

Realizing the vision of using AI agents to automate critical IT tasks depends on the ability to measure and understand effectiveness of proposed solutions. We introduce ITBench, a framework that offers a systematic methodology for benchmarking AI agents to address real-world IT automation tasks. Our…

2021

CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks

NeurIPS 2021poster

Over the last several decades, software has been woven into the fabric of every aspect of our society. As software development surges and code infrastructure of enterprise applications ages, it is now more critical than ever to increase software development productivity and modernize legacy applicat…

Cited by 327SourcecodeScholar
2019

Bias Mitigation Post-processing for Individual and Group Fairness

ICASSP 2019accepted

Whereas previous post-processing approaches for increasing the fairness of predictions of biased classifiers address only group fairness, we propose a method for increasing both individual and group fairness. Our novel framework includes an individual bias detector used to prioritize data samples in…

Cited by 0SourceScholar