← Search

Ruizhi Wang

4 accepted papers

2026

D3-RSMDE: 40× Faster and High-Fidelity Remote Sensing Monocular Depth Estimation

AAAI 2026technical

Real-time, high-fidelity monocular depth estimation from remote sensing imagery is crucial for numerous applications, yet existing methods face a stark trade-off between accuracy and efficiency. Although using Vision Transformer (ViT) backbones for dense prediction is fast, they often exhibit poor p

Cited by 0SourcePDFScholar
2026

Prototype Transformer: Towards Language Model Architectures Interpretable by Design

ICML 2026poster

While state-of-the-art language models (LMs) surpass the vast majority of humans in certain domains, their reasoning remains largely opaque, reducing trust and risking deception and hallucination. In this work, we introduce the Prototype Transformer (ProtoT)—an autoregressive LM architecture that re…

Cited by 0SourceScholar
2023

MPS-AMS: Masked Patches Selection and Adaptive Masking Strategy Based Self-Supervised Medical Image Segmentation

ICASSP 2023accepted

Existing self-supervised learning methods based on contrastive learning and masked image modeling have demonstrated impressive performances. However, current masked image modeling methods are mainly utilized in natural images, and their applications in medical images are relatively lacking. Besides,…

Cited by 0SourceScholar
2023

MvCo-DoT: Multi-View Contrastive Domain Transfer Network for Medical Report Generation

ICASSP 2023accepted

In clinical scenarios, multiple medical images with different views are usually generated at the same time, and they have high semantic consistency. However, the existing medical report generation methods cannot exploit the rich multi-view mutual information of medical images. Therefore, in this wor…

Cited by 0SourceScholar