← Search

Ju Zhang

4 accepted papers

2025

Deformable Attention-Based Edge-Aware Network for Single Image Super-Resolution

ICASSP 2025accepted

Accurately reconstructing object edges is a key challenge in single image super-resolution (SISR), as it greatly influences our visual perception of image quality. To address this fundamental issue, we propose a novel SISR approach named the deformable attention-based edge-aware (DAE) network. The D…

Cited by 0SourceScholar
2022

Using Multiple Reference Audios and Style Embedding Constraints for Speech Synthesis

ICASSP 2022accepted

The end-to-end speech synthesis model can directly take an utterance as reference audio, and generate speech from the text with prosody and speaker characteristics similar to the reference audio. However, an appropriate acoustic embedding must be manually selected during inference. Due to the fact t…

Cited by 0SourceScholar
2021

Improving Naturalness and Controllability of Sequence-to-Sequence Speech Synthesis by Learning Local Prosody Representations

ICASSP 2021accepted

State-of-the-art neural text-to-speech (TTS) networks are trained with a large amount of speech data, which significantly improves the quality of synthetic speech compared with traditional approaches. However, the prosody and controllability of the generated speech is still insufficient, especially…

Cited by 0SourceScholar
2016

Continuous ultrasound based tongue movement video synthesis from speech

ICASSP 2016accepted

The movement of tongue plays an important role in pronunciation. Visualizing the movement of tongue can improve speech intelligibility and also helps learning a second language. However, hardly any research has been investigated for this topic. In this paper, a framework to synthesize continuous ult…

Cited by 0SourceScholar