← Search

Jiaxi Cao

4 accepted papers

2026

E-mem: Multi-Agent Based Episodic Context Reconstruction for LLM Agent Memory

ICML 2026poster

The evolution of Large Language Model (LLM) agents towards System~2 reasoning, characterized by deliberative, high-precision problem-solving, necessitates maintaining rigorous logical integrity over extended horizons. However, prevalent memory preprocessing paradigms incur destructive de-contextuali…

Cited by 0SourceScholar
2026

Fore-Mamba3D: Mamba-based Foreground-Enhanced Encoding for 3D Object Detection

ICLR 2026poster

Linear modeling methods like Mamba have been merged as the effective backbone for the 3D object detection task. However, previous Mamba-based methods utilize the bidirectional encoding for the whole non-empty voxel sequence, which contains abundant useless background information in the scenes. Thoug…

Cited by 0SourcecodeScholar
2026

RoSAMDepth: Robust Self-supervised Depth Estimation Leveraging Segment Anything Model

CVPR 2026

Robust depth estimation aims to maintain high-quality depths across diverse conditions. However, most existing methods estimate depth without taking into account the object-level information. As a result, the predicted depth may easily deviate within objects and become blurred under adverse conditio

Cited by 0SourcecodeScholar
2026

V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning

ICML 2026poster

Multimodal large language models (MLLMs) have achieved remarkable success in general perception, yet complex multi-step visual reasoning remains a persistent challenge. Although recent agentic approaches incorporate tool use, they often neglect critical execution feedback. Consequently, they suffer …

Cited by 0SourceScholar