2026
Measuring Physical-World Privacy Awareness of Large Language Models: An Evaluation Benchmark
ICLR 2026poster
The deployment of Large Language Models (LLMs) in embodied agents creates an urgent need to measure their privacy awareness in the physical world. Existing evaluation methods, however, are confined to natural language based scenarios. To bridge this gap, we introduce EAPrivacy, a comprehensive evalu…