← Search

Boqi Chen

5 accepted papers

2026

YoNoSplat: You Only Need One Model for Feedforward 3D Gaussian Splatting

ICLR 2026poster

Fast and flexible 3D scene reconstruction from unstructured image collections remains a significant challenge. We present YoNoSplat, a feedforward model that reconstructs high-quality 3D Gaussian Splatting representations from an arbitrary number of images. Our model is highly versatile, operating e…

Cited by 0SourceScholar
2025

Seeing Beyond: Enhancing Visual Question Answering with Multi-Modal Retrieval

COLING 2025industry

Multi-modal Large language models (MLLMs) have made significant strides in complex content understanding and reasoning. However, they still suffer from model hallucination and lack of specific knowledge when facing challenging questions. To address these limitations, retrieval augmented generation (…

Cited by 0SourcePDFScholar
2022

Differentiable Zooming for Multiple Instance Learning on Whole-Slide Images

ECCV 2022poster

"Multiple Instance Learning (MIL) methods have become increasingly popular for classifying gigapixel-sized Whole-Slide Images (WSIs) in digital pathology. Most MIL methods operate at a single WSI magnification, by processing all the tissue patches. Such a formulation induces high computational requi…

2021

Detecting Frames in News Headlines and Lead Images in U.S. Gun Violence Coverage

EMNLP 2021finding

News media structure their reporting of events or issues using certain perspectives. When describing an incident involving gun violence, for example, some journalists may focus on mental health or gun regulation, while others may emphasize the discussion of gun rights. Such perspectives are called “…

Cited by 21SourcePDFScholar