← Search

Roman Korostik

3 accepted papers

2025

Generative Speech Foundation Model Pretraining for High-Quality Speech Extraction and Restoration

ICASSP 2025accepted

This paper proposes a generative pretraining foundation model for high-quality speech restoration tasks. By directly operating on complex-valued short-time Fourier transform coefficients, our model does not rely on any vocoders for time-domain signal reconstruction. As a result, our model simplifies…

Cited by 0SourceScholar
2025

Robust Speech Recognition with Schrödinger Bridge-Based Speech Enhancement

ICASSP 2025accepted

In this work, we investigate application of generative speech enhancement to improve the robustness of ASR models in noisy and reverberant conditions. We employ a recently-proposed speech enhancement model based on Schrödinger bridge, which has been shown to perform well compared to diffusion-based…

Cited by 0SourceScholar