← Search

Gayatri Krishnakumar

2 accepted papers

2026

FindTheFlaws: Annotated Errors for Detecting Flawed Reasoning and Scalable Oversight Research

AAAI 2026technical

As AI models tackle increasingly complex problems, ensuring reliable human oversight becomes more challenging due to the difficulty of verifying solutions. Approaches to scaling AI supervision include debate, in which two agents engage in structured dialogue to help a judge evaluate claims; critique

Cited by 0SourcePDFScholar
2025

SandboxSocial: A Sandbox for Social Media Using Multimodal AI Agents

IJCAI 2025

The online information ecosystem enables influence campaigns of unprecedented scale and impact. We urgently need empirically grounded approaches to counter the growing threat of malicious campaigns, now amplified by generative AI. But, developing defenses in real-world settings is impractical. Socia