← Search

Haoran Cheng

8 accepted papers

2026

MCPTox: A Benchmark for Tool Poisoning on Real-World MCP Servers

AAAI 2026technical

By providing a standardized interface for LLM agents to interact with external tools, the Model Context Protocol (MCP) is quickly becoming a cornerstone of the modern autonomous agent ecosystem. However, it creates novel attack surfaces due to untrusted external tools. While prior work has focused o

Cited by 0SourcePDFScholar
2025

Object-level Data Augmentation for Visual 3D Object Detection in Autonomous Driving

ICASSP 2025accepted

Data augmentation plays an important role in visual-based 3D object detection. Existing detectors typically employ image/BEV-level data augmentation techniques, failing to utilize flexible object-level augmentations because of 2D-3D inconsistencies. This limitation hinders us from increasing the div…

Cited by 0SourceScholar
2025

Self-Supervised Direct Preference Optimization for Text-to-Image Diffusion Models

NeurIPS 2025poster

Direct preference optimization (DPO) is an effective method for aligning generative models with human preferences and has been successfully applied to fine‑tune text‑to‑image diffusion models. Its practical adoption, however, is hindered by a labor‑intensive pipeline that first produces a large set…

Cited by 0SourceScholar
2025

VP-MEL: Visual Prompts Guided Multimodal Entity Linking

ACL 2025finding

Multimodal entity linking (MEL), a task aimed at linking mentions within multimodal contexts to their corresponding entities in a knowledge base (KB), has attracted much attention due to its wide applications in recent years. However, existing MEL methods often rely on mention words as retrieval cue…

Cited by 0SourcePDFScholar
2024

Learning Occupancy for Monocular 3D Object Detection

CVPR 2024poster

Monocular 3D detection is a challenging task due to the lack of accurate 3D information. Existing approaches typically rely on geometry constraints and dense depth estimates to facilitate the learning but often fail to fully exploit the benefits of three-dimensional feature extraction in frustum and…

2024

Regulating Intermediate 3D Features for Vision-Centric Autonomous Driving

AAAI 2024technical

Multi-camera perception tasks have gained significant attention in the field of autonomous driving. However, existing frameworks based on Lift-Splat-Shoot (LSS) in the multi-camera setting cannot produce suitable dense 3D features due to the projection nature and uncontrollable densification process…

2023

Coupled Point Process-based Sequence Modeling for Privacy-preserving Network Alignment

IJCAI 2023poster

Network alignment aims at finding the correspondence of nodes across different networks, which is significant for many applications, e.g., fraud detection and crime network tracing across platforms. In practice, however, accessing the topological information of different networks is often restrict…

2023

MonoNeRD: NeRF-like Representations for Monocular 3D Object Detection

ICCV 2023poster

In the field of monocular 3D detection, it is common practice to utilize scene geometric clues to enhance the detector's performance. However, many existing works adopt these clues explicitly such as estimating a depth map and back-projecting it into 3D space. This explicit methodology induces spars…

Cited by 35PDFcodeScholar