← Search

Linda Shapiro

10 accepted papers

2025

BADGR: Bundle Adjustment Diffusion Conditioned by Gradients for Wide-Baseline Floor Plan Reconstruction

CVPR 2025highlight

Reconstructing precise camera poses and floor plan layouts from wide-baseline RGB panoramas is a difficult and unsolved problem. We introduce BADGR, a novel diffusion model that jointly performs reconstruction and bundle adjustment (BA) to refine poses and layouts from a coarse state, using 1D floor…

Cited by 0SourcePDFScholar
2025

PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology

ICCV 2025poster

Diagnosing diseases through histopathology whole slide images (WSIs) is fundamental in modern pathology but is challenged by the gigapixel scale and complexity of WSIs. Trained histopathologists overcome this challenge by navigating the WSI, looking for relevant patches, taking notes, and compiling…

Cited by 0SourcePDFScholar
2024

Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos

CVPR 2024poster

Diagnosis in histopathology requires a global whole slide images (WSIs) analysis requiring pathologists to compound evidence from different WSI patches. The gigapixel scale of WSIs poses a challenge for histopathology multi-modal models. Training multi-model models for histopathology requires instru…

Cited by 34SourcePDFScholar
2023

Quilt-1M: One Million Image-Text Pairs for Histopathology

NeurIPS 2023oral

Recent accelerations in multi-modal applications have been made possible with the plethora of image and text data available online. However, the scarcity of analogous data in the medical field, specifically in histopathology, has slowed comparable progress. To enable similar representation learning…

2021

Semi-Supervised Synthesis of High-Resolution Editable Textures for 3D Humans

CVPR 2021poster

We introduce a novel approach to generate diverse high fidelity texture maps for 3D human meshes in a semi-supervised setup. Given a segmentation mask defining the layout of the semantic regions in the texture map, our network generates high-resolution textures with a variety of styles, that are the…

Cited by 24PDFScholar
2020

Personalized Face Modeling for Improved Face Reconstruction and Motion Retargeting

ECCV 2020poster

Traditional methods for image-based 3D face reconstruction and facial motion retargeting fit a 3D morphable model (3DMM) to the face, which has limited modeling capacity and fail to generalize well to in-the-wild data. Use of deformation transfer or multilinear tensor as a personalized 3DMM for blen…

Cited by 74SourcePDFScholar
2019

ESPNetv2: A Light-Weight, Power Efficient, and General Purpose Convolutional Neural Network

CVPR 2019poster

We introduce a light-weight, power efficient, and general purpose convolutional neural network, ESPNetv2, for modeling visual and sequential data. Our network uses group point-wise and depth-wise dilated separable convolutions to learn representations from a large effective receptive field with fewe…

Cited by 600PDFcodeScholar
2018

ESPNet: Efficient Spatial Pyramid of Dilated Convolutions for Semantic Segmentation

ECCV 2018poster

We introduce a fast and efficient convolutional neural network, ESPNet, for semantic segmentation of high resolution images under resource constraints. ESPNet is based on a new convolutional module, efficient spatial pyramid (ESP), which is efficient in terms of computation, memory, and power. ESPNe…

2015

Generating Notifications for Missing Actions: Don't Forget to Turn the Lights Off!

ICCV 2015poster

We all have experienced forgetting habitual actions among our daily activities. For example, we probably have forgotten to turn the lights off before leaving a room or turn the stove off after cooking. In this paper, we propose a solution to the problem of issuing notifications on actions that may b…

Cited by 85PDFScholar