← Search

Ruiying Lu

9 accepted papers

2024

Latent Diffusion Prior Enhanced Deep Unfolding for Snapshot Spectral Compressive Imaging

ECCV 2024oral

"Snapshot compressive spectral imaging reconstruction aims to reconstruct three-dimensional spatial-spectral images from a single-shot two-dimensional compressed measurement. Existing state-of-the-art methods are mostly based on deep unfolding structures but have intrinsic performance bottlenecks: i…

2023

ConZIC: Controllable Zero-Shot Image Captioning by Sampling-Based Polishing

CVPR 2023poster

Zero-shot capability has been considered as a new revolution of deep learning, letting machines work on tasks without curated training data. As a good start and the only existing outcome of zero-shot image captioning (IC), ZeroCap abandons supervised training and sequentially searching every word in…

2023

Hierarchical Vector Quantized Transformer for Multi-class Unsupervised Anomaly Detection

NeurIPS 2023poster

Unsupervised image Anomaly Detection (UAD) aims to learn robust and discriminative representations of normal samples. While separate solutions per class endow expensive computation and limited generalizability, this paper focuses on building a unified framework for multiple classes. Under such a cha…

2023

PatchCT: Aligning Patch Set and Label Set with Conditional Transport for Multi-Label Image Classification

ICCV 2023poster

Multi-label image classification is a prediction task that aims to identify more than one label from a given image. This paper considers the semantic consistency of the latent space between the visual patch and linguistic label domains and introduces the conditional transport (CT) theory to bridge t…

Cited by 22PDFcodeScholar
2022

HyperMiner: Topic Taxonomy Mining with Hyperbolic Embedding

NeurIPS 2022accept

Embedded topic models are able to learn interpretable topics even with large and heavy-tailed vocabularies. However, they generally hold the Euclidean embedding space assumption, leading to a basic limitation in capturing hierarchical relations. To this end, we present a novel framework that introdu…

2021

Memory-Efficient Network for Large-Scale Video Compressive Sensing

CVPR 2021poster

Video snapshot compressive imaging (SCI) captures a sequence of video frames in a single shot using a 2D detector. The underlying principle is that during one exposure time, different masks are imposed on the high-speed scene to form a compressed measurement. With the knowledge of masks, optimizatio…

Cited by 93PDFcodeScholar
2020

BIRNAT: Bidirectional Recurrent Neural Networks with Adversarial Training for Video Snapshot Compressive Imaging

ECCV 2020poster

We consider the problem of video snapshot compressive imaging (SCI), where multiple high-speed frames are coded by different masks and then summed to a single measurement. This measurement and the modulation masks are fed into our Recurrent Neural Network (RNN) to reconstruct the desired high-speed…

2020

Recurrent Hierarchical Topic-Guided RNN for Language Generation

ICML 2020poster

To simultaneously capture syntax and global semantics from a text corpus, we propose a new larger-context recurrent neural network (RNN) based language model, which extracts recurrent hierarchical semantic structure via a dynamic deep topic model to guide natural language generation. Moving beyond a…