← Search

Behzad Bozorgtabar

17 accepted papers

2026

Does a Hybrid Space-Aware Randomized Defense Improve Empirical and Certified Adversarial Robustness?

ICML 2026poster

We introduce Hybrid Space-aware Stochastic Convolution Attention Noise (HySCAN), a hybrid randomized defense that helps close the long-standing gap between provable robustness under ℓ2 certificates and empirical robustness against strong ℓ∞ attacks, while maintaining strong generalization across div…

Cited by 0SourceScholar
2026

LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs

CVPR 2026

Medical vision-language models (VLMs) are strong zero-shot recognizers for medical imaging, but their reliability under domain shift hinges on calibrated uncertainty with guarantees. Split conformal prediction (SCP) offers finite-sample coverage, yet prediction sets often become large (low efficienc

Cited by 0SourceScholar
2026

VALIANT: Prompt Instability for Active Learning in Black-Box Medical Imaging

AAAI 2026technical

The deployment of large, black-box foundation models for medical image classification is often hindered by the high cost of acquiring large, task-specific labeled datasets for fine-tuning. While active learning (AL) presents a promising solution, many state-of-the-art AL methods are computationally

Cited by 0SourcePDFScholar
2025

A Simple Framework for Open-Vocabulary Zero-Shot Segmentation

ICLR 2025poster

Zero-shot classification capabilities naturally arise in models trained within a vision-language contrastive framework. Despite their classification prowess, these models struggle in dense tasks like zero-shot open-vocabulary segmentation. This deficiency is often attributed to the absence of locali…

2025

ReservoirTTA: Prolonged Test-time Adaptation for Evolving and Recurring Domains

NeurIPS 2025poster

This paper introduces **ReservoirTTA**, a novel plug–in framework designed for prolonged test–time adaptation (TTA) in scenarios where the test domain continuously shifts over time, including cases where domains recur or evolve gradually. At its core, ReservoirTTA maintains a reservoir of domain-spe…

Cited by 0SourcecodeScholar
2025

UniViT: Unifying Image and Video Understanding in One Vision Encoder

NeurIPS 2025poster

Despite the impressive progress of recent pretraining methods on multimodal tasks, existing methods are inherently biased towards either spatial modeling (e.g., CLIP) or temporal modeling (e.g., V-JEPA), limiting their joint capture of spatial details and temporal dynamics. To this end, we propose U…

Cited by 0SourceScholar
2024

Combining Graph Transformers Based Multi-Label Active Learning and Informative Data Augmentation for Chest Xray Classification

AAAI 2024technical

Informative sample selection in active learning (AL) helps a machine learning system attain optimum performance with minimum labeled samples, thus improving human-in-the-loop computer-aided diagnosis systems with limited labeled data. Data augmentation is highly effective for enlarging datasets with…

Cited by 1SourcePDFScholar
2024

CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping

ICLR 2024spotlight

Leveraging nearest neighbor retrieval for self-supervised representation learning has proven beneficial with object-centric images. However, this approach faces limitations when applied to scene-centric datasets, where multiple objects within an image are only implicitly captured in the global repre…

2024

Un-Mixing Test-Time Normalization Statistics: Combatting Label Temporal Correlation

ICLR 2024poster

Recent test-time adaptation methods heavily rely on nuanced adjustments of batch normalization (BN) parameters. However, one critical assumption often goes overlooked: that of independently and identically distributed (i.i.d.) test batches with respect to unknown labels. This oversight leads to ske…

2023

Adaptive Similarity Bootstrapping for Self-Distillation Based Representation Learning

ICCV 2023poster

Most self-supervised methods for representation learning leverage a cross-view consistency objective i.e., they maximize the representation similarity of a given image's augmented views. Recent work NNCLR goes beyond the cross-view paradigm and uses positive pairs from different images obtained via…

Cited by 2PDFcodeScholar
2023

Attention-Conditioned Augmentations for Self-Supervised Anomaly Detection and Localization

AAAI 2023technical

Self-supervised anomaly detection and localization are critical to real-world scenarios in which collecting anomalous samples and pixel-wise labeling is tedious or infeasible, even worse when a wide variety of unseen anomalies could surface at test time. Our approach involves a pretext task in the c…

Cited by 22SourcePDFScholar
2023

CrOC: Cross-View Online Clustering for Dense Visual Representation Learning

CVPR 2023poster

Learning dense visual representations without labels is an arduous task and more so from scene-centric data. We propose to tackle this challenging problem by proposing a Cross-view consistency objective with an Online Clustering mechanism (CrOC) to discover and segment the semantics of the views. In…

2023

TeSLA: Test-Time Self-Learning With Automatic Adversarial Augmentation

CVPR 2023poster

Most recent test-time adaptation methods focus on only classification tasks, use specialized network architectures, destroy model calibration or rely on lightweight information from the source domain. To tackle these issues, this paper proposes a novel Test-time Self-Learning method with automatic A…

2021

Quantifying Explainers of Graph Neural Networks in Computational Pathology

CVPR 2021poster

Explainability of deep learning methods is imperative to facilitate their clinical adoption in digital pathology. However, popular deep learning methods and explainability techniques (explainers) based on pixel-wise processing disregard biological entities' notion, thus complicating comprehension by…

Cited by 107PDFcodeScholar
2020

Pathological Retinal Region Segmentation From OCT Images Using Geometric Relation Based Augmentation

CVPR 2020poster

Medical image segmentation is important for computer aided diagnosis. Pixelwise manual annotations of large datasets require high expertise and is time consuming. Conventional data augmentations have limited benefit by not fully representing the underlying distribution of the training set, thus affe…

Cited by 46PDFScholar
2019

SROBB: Targeted Perceptual Loss for Single Image Super-Resolution

ICCV 2019poster

By benefiting from perceptual losses, recent studies have improved significantly the performance of the super-resolution task, where a high-resolution image is resolved from its low-resolution counterpart. Although such objective functions generate near-photorealistic results, their capability is li…

Cited by 180PDFcodeScholar
2019

SynDeMo: Synergistic Deep Feature Alignment for Joint Learning of Depth and Ego-Motion

ICCV 2019poster

Despite well-established baselines, learning of scene depth and ego-motion from monocular video remains an ongoing challenge, specifically when handling scaling ambiguity issues and depth inconsistencies in image sequences. Much prior work uses either a supervised mode of learning or stereo images.…

Cited by 47PDFScholar