← Search

Guillermo Cámbara

3 accepted papers

2024

Mapache: Masked Parallel Transformer for Advanced Speech Editing and Synthesis

ICASSP 2024accepted

Recent advancements in Generative AI, such as scaled Transformer large language models (LLM) and diffusion decoders, have revolutionized speech synthesis. With speech encompassing the complexities of natural language and audio dimensionality, many recent models have relied on autoregressive modeling…

Cited by 0SourceScholar
2022

Recycle Your Wav2Vec2 Codebook: A Speech Perceiver for Keyword Spotting

COLING 2022main

Speech information in a pretrained wav2vec2.0 model is usually leveraged through its encoder, which has at least 95M parameters, being not so suitable for small footprint Keyword Spotting. In this work, we show an efficient way of profiting from wav2vec2.0’s linguistic knowledge, by recycling the ph…

2020

Detection of Speech Events and Speaker Characteristics through Photo-Plethysmographic Signal Neural Processing

ICASSP 2020accepted

The use of photoplethysmogram signal (PPG) for heart and sleep monitoring is commonly found nowadays in smart-phones and wrist wearables. Besides common usages, it has been proposed and reported that person information can be extracted from PPG for other uses, like biometry tasks. In this work, we e…

Cited by 0SourceScholar