← Search

Michael Green

4 accepted papers

2026

Fast Autoregressive Video Diffusion and World Models with Temporal Cache Compression and Sparse Attention

ICML 2026poster

Autoregressive video diffusion models enable \emph{streaming} generation, opening the door to long-form synthesis, video world models, and interactive neural game engines. However, their core attention layers become a major bottleneck at inference time: as generation progresses, the KV cache grows, …

Cited by 4SourceScholar
2025

EffoVPR: Effective Foundation Model Utilization for Visual Place Recognition

ICLR 2025poster

The task of Visual Place Recognition (VPR) is to predict the location of a query image from a database of geo-tagged images. Recent studies in VPR have highlighted the significant advantage of employing pre-trained foundation models like DINOv2 for the VPR task. However, these models are often deeme…

Cited by 10SourcePDFScholar
2025

Find your Needle: Small Object Image Retrieval via Multi-Object Attention Optimization

NeurIPS 2025poster

We address the challenge of Small Object Image Retrieval (SoIR), where the goal is to retrieve images containing a specific small object, in a cluttered scene. The key challenge in this setting is constructing a single image descriptor, for scalable and efficient search, that effectively represents…

Cited by 0SourceScholar
2023

Automatic Error Detection in Integrated Circuits Image Segmentation: A Data-Driven Approach

ICASSP 2023accepted

Due to the complicated nanoscale structures of current integrated circuits(IC) builds and low error tolerance of IC image segmentation tasks, most existing automated IC image segmentation approaches require human experts for visual inspection to ensure correctness, which is one of the major bottlene…

Cited by 0SourceScholar