← Search

Jian Xiong

9 accepted papers

2026

Point Cloud Quality Assessment via Multi-View Structure-Aware Feature Fusion

AAAI 2026technical

Point cloud quality assessment (PCQA) is essential for reliable 3D visual applications. While point-based methods face challenges in characterizing distortions due to point cloud disorder, projection-based approaches offer better efficiency but suffer from geometric distortion insensitivity and text

Cited by 0SourcePDFScholar
2026

UGround: Towards Unified Visual Grounding with Unrolled Transformers

ICML 2026poster

We present UGround, a **U**nified visual **Ground**ing paradigm that dynamically selects intermediate layers across **U**nrolled transformers as "mask as prompt'', diverging from the prevailing pipeline that leverages the fixed last hidden layer as "$\texttt{\}$ as prompt''. UGround addresses two pr…

Cited by 0SourceScholar
2025

Adaptive Skeleton Prompt Tuning for Cross-Dataset 3D Human Pose Estimation

ICASSP 2025accepted

Inconsistency of distributions in human actions and camera viewpoints can lead to significant deviations when the pre-trained 3D pose estimators are tested on cross-datasets. In practical applications, the estimators usually follow the standard full fine-tuning paradigm on the target dataset, which…

Cited by 0SourceScholar
2025

RFEM: Remote Feature Enhancement Module for Target Detection

ICASSP 2025accepted

The research and development of dense crowd detection technology have always been one of the hot and challenging topics in the field of computer vision. DETR-like models have shown good performance in both training efficiency and inference capabilities. Nevertheless, as the optimization proceeds, th…

Cited by 0SourceScholar
2024

DeformMLP: Dynamic Large-Scale Receptive Field MLP Networks for Human Motion Prediction

ICASSP 2024accepted

Predicting human motion requires addressing dependencies and errors for pose forecasting from sequences. The transformer’s self-attention aids this, but its complexity poses computational challenges. We present an efficient DeformMLP network without self-attention, using fully connected layers. Defo…

Cited by 0SourceScholar
2024

Geometry Compression Artifact Removal for V-PCC over a Wide Bitrate Range

ICASSP 2024accepted

In video-based point cloud compression (V-PCC), point clouds are generated as videos via patch projection to be compressed using video coding techniques. However, a large number of filled empty pixels in the videos creates a fake context, which reduces the noise prediction accuracy in compression ar…

Cited by 0SourceScholar
2023

Learning Hybrid Representations of Semantics and Distortion for Blind Image Quality Assessment

ICASSP 2023accepted

Recently, some studies have shown that semantic and distortion representations both benefit the evaluation of image quality. However, the images of existing synthetic distortion databases are annotated with subjective quality scores and distortion types, lacking labels with semantic objects. Therefo…

Cited by 0SourceScholar
2023

ψ-Net: Point Structural Information Network for No-Reference Point Cloud Quality Assessment

ICASSP 2023accepted

The human vision system is highly adapted to extract structural information from the viewed scenes. The irregularity of point clouds makes the extraction of structural information containing both color and geometry an important challenge for point cloud quality assessment (PCQA). This paper proposes…

Cited by 0SourceScholar