← Search

Wenrui Li

11 accepted papers

2026

DYNAMICAL ISOMETRY BASED RIGOROUS FAIR NEURAL ARCHITECTURE SEARCH

ICASSP 2026poster

Recently, the weight-sharing technique has significantly speeded up the training and evaluation procedure of neural architecture search. However, most existing weight-sharing strategies are solely based on experience or observation, which makes the searching results lack interpretability and rationa…

Cited by 0SourcePDFScholar
2026

Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D Retrieval

AAAI 2026technical

With the daily influx of 3D data on the internet, text-3D retrieval has gained increasing attention. However, current methods face two major challenges: Hierarchy Representation Collapse (HRC) and Redundancy-Induced Saliency Dilution (RISD). HRC compresses abstract-to-specific and whole-to-part hier

Cited by 0SourcePDFScholar
2026

MRT: Learning Compact Representations with Mixed RWKV-Transformer for Extreme Image Compression

AAAI 2026technical

Recent advances in extreme image compression have revealed that mapping pixel data into highly compact latent representations can significantly improve coding efficiency. However, most existing methods compress images into 2-D latent spaces via convolutional neural networks (CNNs) or Swin Transforme

Cited by 0SourcePDFScholar
2026

T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates

AAAI 2026technical

Recent advances in video generation techniques have given rise to an emerging paradigm of generative video coding for Ultra-Low Bitrate (ULB) scenarios by leveraging powerful generative priors. However, most existing methods are limited by domain specificity (e.g., facial or human videos) or excessi

Cited by 0SourcePDFScholar
2025

Digging into Intrinsic Contextual Information for High-fidelity 3D Point Cloud Completion

AAAI 2025technical

The common occurrence of occlusion-induced incompleteness in point clouds has made point cloud completion (PCC) a highly-concerned task in the field of geometric processing. Existing PCC methods typically produce complete point clouds from partial point clouds in a coarse-to-fine paradigm, with the…

2025

Hyperbolic-Constraint Point Cloud Reconstruction from Single RGB-D Images

AAAI 2025technical

Reconstructing desired objects and scenes has long been a primary goal in 3D computer vision. Single-view point cloud reconstruction has become a popular technique due to its low cost and accurate results. However, single-view reconstruction methods often rely on expensive CAD models and complex geo…

Cited by 0SourcePDFScholar
2025

Riemann-based Multi-scale Attention Reasoning Network for Text-3D Retrieval

AAAI 2025technical

Due to the challenges in acquiring paired Text-3D data and the inherent irregularity of 3D data structures, combined representation learning of 3D point clouds and text remains unexplored. In this paper, we propose a novel Riemann-based Multi-scale Attention Reasoning Network (RMARN) for text-3D ret…

2025

Text-Guided Editable 3D City Scene Generation

ICASSP 2025accepted

The automated generation of 3D city scenes has attracted considerable attention due to its broad applications in areas such as virtual reality, urban planning, and digital media. Traditional approaches for constructing 3D city environments typically depend on labor-intensive manual modeling or the u…

Cited by 0SourceScholar
2024

MIntRec2.0: A Large-scale Benchmark Dataset for Multimodal Intent Recognition and Out-of-scope Detection in Conversations

ICLR 2024poster

Multimodal intent recognition poses significant challenges, requiring the incorporation of non-verbal modalities from real-world contexts to enhance the comprehension of human intentions. However, most existing multimodal intent benchmark datasets are limited in scale and suffer from difficulties in…