2026
LLM Safety in Judicial AI: A Stress Test of Social Media Influence on Real-World Judgments
AAAI 2026technical
Integrating Large Language Models (LLMs) into judicial decision-making demands rigorous safety examination against non-legal influences. This paper presents a novel stress test where we evaluate LLM-generated labor dispute outcomes by introducing social media sentiment as an external pressure, criti