← Search

Lucas Ventura

3 accepted papers

2025

Chapter-Llama: Efficient Chaptering in Hour-Long Videos with LLMs

CVPR 2025poster

We address the task of video chaptering, i.e., partitioning a long video timeline into semantic units and generating corresponding chapter titles. While relatively underexplored, automatic chaptering has the potential to enable efficient navigation and content retrieval in long-form videos. In this…

Cited by 0SourcePDFScholar
2024

CoVR: Learning Composed Video Retrieval from Web Video Captions

AAAI 2024technical

Composed Image Retrieval (CoIR) has recently gained popularity as a task that considers both text and image queries together, to search for relevant images in a database. Most CoIR approaches require manually annotated datasets, comprising image-text-image triplets, where the text describes a modifi…

Cited by 46SourcePDFScholar
2021

How2Sign: A Large-Scale Multimodal Dataset for Continuous American Sign Language

CVPR 2021poster

One of the factors that have hindered progress in the areas of sign language recognition, translation, and production is the absence of large annotated datasets. Towards this end, we introduce How2Sign, a multimodal and multiview continuous American Sign Language (ASL) dataset, consisting of a paral…

Cited by 257PDFcodeScholar