← Search

Aimon Rahman

2 accepted papers

2026

MOVi: Training-free Text-conditioned Multi-Object Video Generation

ICASSP 2026oral

Recent advances in diffusion-based text-to-video (T2V) models have demonstrated remarkable progress, but these models still face challenges in generating videos with multiple objects. Most models struggle with accurately capturing complex object interactions, often treating some objects as static ba…

Cited by 0SourcePDFScholar
2023

Ambiguous Medical Image Segmentation Using Diffusion Models

CVPR 2023poster

Collective insights from a group of experts have always proven to outperform an individual's best diagnostic for clinical tasks. For the task of medical image segmentation, existing research on AI-based alternatives focuses more on developing models that can imitate the best individual rather than h…