← Search

Liyang Lu

3 accepted papers

2023

i-Code: An Integrative and Composable Multimodal Learning Framework

AAAI 2023technical

Human intelligence is multimodal; we integrate visual, linguistic, and acoustic signals to maintain a holistic worldview. Most current pretraining methods, however, are limited to one or two modalities. We present i-Code, a self-supervised pretraining framework where users may flexibly combine the m…

2021

Generating Human Readable Transcript for Automatic Speech Recognition with Pre-Trained Language Model

ICASSP 2021accepted

Modern Automatic Speech Recognition (ASR) systems can achieve high performance in terms of recognition accuracy. However, a perfectly accurate transcript still can be challenging to read due to disfluency, filter words, and other errata common in spoken communication. Many downstream tasks and human…

Cited by 0SourceScholar
2015

Single-Shot Specular Surface Reconstruction With Gonio-Plenoptic Imaging

ICCV 2015poster

We present a gonio-plenoptic imaging system that realizes a single-shot shape measurement for specular surfaces. The system is comprised of a collimated illumination source and a plenoptic camera. Unlike a conventional plenoptic camera, our system captures the BRDF variation of the object surface in…

Cited by 11PDFScholar