← Search

Emanuel Ben-Baruch

6 accepted papers

2026

Scene-VLM: Multimodal Video Scene Segmentation via Vision-Language Models

CVPR 2026

Segmenting long-form videos into semantically coherent scenes is a fundamental task in large-scale video understanding. Existing encoder-based methods are limited by visual-centric biases, classify each shot in isolation without leveraging sequential dependencies, and lack both narrative understandi

Cited by 0SourceScholar
2025

LV-MAE: Learning Long Video Representations through Masked-Embedding Autoencoders

ICCV 2025poster

In this work, we introduce long-video masked-embedding autoencoders (LV-MAE), a self-supervised learning framework for long video representation.Our approach treats short- and long-span dependencies as two separate tasks.Such decoupling allows for a more intuitive video processing where short-span s…

Cited by 0SourcePDFScholar
2022

Multi-Label Classification With Partial Annotations Using Class-Aware Selective Loss

CVPR 2022poster

Large-scale multi-label classification datasets are commonly, and perhaps inevitably, partially annotated. That is, only a small subset of labels are annotated per sample. Different methods for handling the missing labels induce different properties on the model and impact its accuracy. In this work…

Cited by 55PDFcodeScholar
2021

Asymmetric Loss for Multi-Label Classification

ICCV 2021poster

In a typical multi-label setting, a picture contains on average few positive labels, and many negative ones. This positive-negative imbalance dominates the optimization process, and can lead to under-emphasizing gradients from positive labels during training, resulting in poor accuracy. In this pape…

Cited by 547PDFcodeScholar
2021

Semantic Diversity Learning for Zero-Shot Multi-Label Classification

ICCV 2021poster

Training a neural network model for recognizing multiple labels associated with an image, including identifying unseen labels, is challenging, especially for images that portray numerous semantically diverse labels. As challenging as this task is, it is an essential task to tackle since it represent…

Cited by 46PDFcodeScholar