← Search

Srikar Yellapragada

5 accepted papers

2025

GECKO: Gigapixel Vision-Concept Contrastive Pretraining in Histopathology

ICCV 2025poster

Pretraining a Multiple Instance Learning (MIL) aggregator enables the derivation of Whole Slide Image (WSI)-level embeddings from patch-level representations without supervision. While recent multimodal MIL pretraining approaches leveraging auxiliary modalities have demonstrated performance gains ov…

2025

Leveraging Registers in Vision Transformers for Robust Adaptation

ICASSP 2025accepted

Vision Transformers (ViTs) have shown success across a variety of tasks due to their ability to capture global image representations. Recent studies have identified the existence of high-norm tokens in ViTs, which can interfere with unsupervised object discovery. To address this, the use of "registe…

Cited by 3SourceScholar
2025

ZoomLDM: Latent Diffusion Model for Multi-scale Image Generation

CVPR 2025poster

Diffusion models have revolutionized image generation, yet several challenges restrict their application to large-image domains, such as digital pathology and satellite imagery. Given that it is infeasible to directly train a model on 'whole' images from domains with potential gigapixel sizes, diffu…

2024

Learned Representation-Guided Diffusion Models for Large-Image Generation

CVPR 2024poster

To synthesize high-fidelity samples diffusion models typically require auxiliary data to guide the generation process. However it is impractical to procure the painstaking patch-level annotation effort required in specialized domains like histopathology and satellite imagery; it is often performed b…

2024

∞-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions

ECCV 2024poster

"Synthesizing high-resolution images from intricate, domain-specific information remains a significant challenge in generative modeling, particularly for applications in large-image domains such as digital histopathology and remote sensing. Existing methods face critical limitations: conditional dif…