← Search

Sarthak Bansal

1 accepted papers

2025

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues

ICASSP 2025accepted

This paper introduces ViNet-S, a 36MB model based on the ViNet architecture with a U-Net design, featuring a lightweight decoder that significantly reduces model size and parameters without compromising performance. Additionally, ViNet-A (148MB) incorporates spatio-temporal action localization (STAL…

Cited by 0SourceScholar