← Search

Wele Gedara Chaminda Bandara

6 accepted papers

2025

Stable Diffusion Models are Secretly Good at Visual In-Context Learning

ICCV 2025poster

Large language models (LLM) in natural language processing (NLP) have demonstrated great potential for in-context learning (ICL) -- the ability to leverage a few set of example prompts to adapt to various tasks without having to explicitly update model weights. ICL has recently been explored for the…

Cited by 0SourcePDFScholar
2024

CrowdDiff: Multi-hypothesis Crowd Density Estimation using Diffusion Models

CVPR 2024poster

Crowd counting is a fundamental problem in crowd analysis which is typically accomplished by estimating a crowd density map and summing over the density values. However this approach suffers from background noise accumulation and loss of density due to the use of broad Gaussian kernels to create the…

2023

AdaMAE: Adaptive Masking for Efficient Spatiotemporal Learning With Masked Autoencoders

CVPR 2023poster

Masked Autoencoders (MAEs) learn generalizable representations for image, text, audio, video, etc., by reconstructing masked input data from tokens of the visible data. Current MAE approaches for videos rely on random patch, tube, or frame based masking strategies to select these tokens. This paper…

2023

Unite and Conquer: Plug & Play Multi-Modal Synthesis Using Diffusion Models

CVPR 2023poster

Generating photos satisfying multiple constraints finds broad utility in the content creation industry. A key hurdle to accomplishing this task is the need for paired data consisting of all modalities (i.e., constraints) and their corresponding output. Moreover, existing methods need retraining usin…

2022

HyperTransformer: A Textural and Spectral Feature Fusion Transformer for Pansharpening

CVPR 2022poster

Pansharpening aims to fuse a registered high-resolution panchromatic image (PAN) with a low-resolution hyperspectral image (LR-HSI) to generate an enhanced HSI with high spectral and spatial resolution. Existing pansharpening approaches neglect using an attention mechanism to transfer HR texture fea…

Cited by 144PDFcodeScholar
2022

SPIN Road Mapper: Extracting Roads from Aerial Images via Spatial and Interaction Space Graph Reasoning for Autonomous Driving

ICRA 2022poster

Road extraction is an essential step in building autonomous navigation systems. Detecting road segments is challenging as they are of varying widths, bifurcated throughout the image, and are often occluded by terrain, cloud, or other weather conditions. Using just convolution neural networks (ConvNe…

Cited by 50SourcecodeScholar