← Search

Roi Benita

4 accepted papers

2026

Joint Enhancement and Classification using Coupled Diffusion Models of Signals and Logits

ICML 2026poster

Robust classification in noisy environments remains a fundamental challenge in machine learning. Standard approaches typically treat signal enhancement and classification as separate, sequential stages: first enhancing the signal and then applying a classifier. This approach fails to leverage the se…

Cited by 0SourceScholar
2025

CAFA: a Controllable Automatic Foley Artist

ICCV 2025poster

Foley is a key element in video production, refers to the process of adding an audio signal to a silent video while ensuring semantic and temporal alignment. In recent years, the rise of personalized content creation and advancements in automatic video-to-audio models have increased the demand for g…

2024

DiffAR: Denoising Diffusion Autoregressive Model for Raw Speech Waveform Generation

ICLR 2024poster

Diffusion models have recently been shown to be relevant for high-quality speech generation. Most work has been focused on generating spectrograms, and as such, they further require a subsequent model to convert the spectrogram to a waveform (i.e., a vocoder). This work proposes a diffusion probabil…