← Search

Tongxin Yuan

4 accepted papers

2025

Caution for the Environment: Multimodal LLM Agents are Susceptible to Environmental Distractions

ACL 2025long

This paper investigates the faithfulness of multimodal large language model (MLLM) agents in a graphical user interface (GUI) environment, aiming to address the research question of whether multimodal GUI agents can be distracted by environmental context. A general scenario is proposed where both th…

2025

Point Cloud Registration via Reconstruction with Local Geometry Information Aggregation

ICASSP 2025accepted

Point cloud registration is a fundamental yet challenging task in computer vision and robotics. While framing it as a reconstruction problem has shown promise, traditional reconstruction approaches rely on positional encodings to encode positional information, which inadequately capture the intricat…

Cited by 0SourceScholar
2024

R-Judge: Benchmarking Safety Risk Awareness for LLM Agents

EMNLP 2024finding

Large language models (LLMs) have exhibited great potential in autonomously completing tasks across real-world applications. Despite this, these LLM agents introduce unexpected safety risks when operating in interactive environments. Instead of centering on the harmlessness of LLM-generated content…