← Search

Jinseok Park

4 accepted papers

2025

Accelerating Codec-based Speech Synthesis with Multi-Token Prediction and Speculative Decoding

ICASSP 2025accepted

The goal of this paper is to accelerate codec-based speech synthesis systems with minimum sacrifice to speech quality. We propose an enhanced inference method that allows for flexible trade-offs between speed and quality during inference without requiring additional training. Our core idea is to pre…

Cited by 0SourceScholar
2024

Learning Contextualized Representation on Discrete Space Via Hierarchical Product Quantization

ICASSP 2024accepted

Self-supervised learning has recently demonstrated significant success in various speech processing applications. Recent studies report that pre-training with contextualized continuous targets plays a crucial role in fine-tuning for better speech downstream tasks. However, unlike the continuous targ…

Cited by 0SourceScholar
2023

Joint Unsupervised and Supervised Learning for Context-Aware Language Identification

ICASSP 2023accepted

Language identification (LID) recognizes the language of a spoken utterance automatically. According to recent studies, LID models trained with an automatic speech recognition (ASR) task perform better than those trained with a LID task only. However, we need additional text labels to train the mode…

Cited by 0SourceScholar
2018

Double JPEG Detection in Mixed JPEG Quality Factors using Deep Convolutional Neural Network

ECCV 2018poster

Double JPEG detection is essential for detecting various image manipulations. This paper proposes a novel deep convolutional neural network for double JPEG detection using statistical histogram features from each block with a vectorized quantization table. In contrast to previous methods, the propos…

Cited by 108SourcePDFScholar