AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
The integration of vision-language models (VLMs) is driving a new generation of embodied agents capable of operating in human-centered environments. However, as deployment expands, these systems face growing safety risks, particularly when executing hazardous instructions. Current safety evaluation