← Search

Zun Liu

3 accepted papers

2026

UNeMo: Collaborative Visual-Language Reasoning and Navigation via a Multimodal World Model

AAAI 2026technical

Vision-and-Language Navigation (VLN) requires agents to autonomously navigate complex environments via visual images and natural language instructions—remains highly challenging. Recent research on enhancing language-guided navigation reasoning using pre-trained large language models (LLMs) has show

Cited by 0SourcePDFScholar
2023

The Devil is in the Crack Orientation: A New Perspective for Crack Detection

ICCV 2023poster

Cracks are usually curve-like structures that are the focus of many computer-vision applications (e.g., road safety inspection and surface inspection of industrial facilities). The existing pixel-based crack segmentation methods rely on time-consuming and costly pixel-level annotations. And the obje…

Cited by 24PDFScholar
2022

Geometry-Aware Guided Loss for Deep Crack Recognition

CVPR 2022poster

Despite the substantial progress of deep models for crack recognition, due to the inconsistent cracks in varying sizes, shapes, and noisy background textures, there still lacks the discriminative power of the deeply learned features when supervised by the cross-entropy loss. In this paper, we propos…

Cited by 35PDFScholar