← Search

Anh Thai

6 accepted papers

2025

SplatTalk: 3D VQA with Gaussian Splatting

ICCV 2025poster

Language-guided 3D scene understanding is important for advancing applications in robotics, AR/VR, and human-computer interaction, enabling models to comprehend and interact with 3D environments through natural language. While 2D vision-language models (VLMs) have achieved remarkable success in 2D V…

Cited by 0SourcePDFScholar
2025

Symmetry Strikes Back: From Single-Image Symmetry Detection to 3D Generation

CVPR 2025highlight

Symmetry is a ubiquitous and fundamental property in the visual world, serving as a critical cue for perception and structure interpretation. This paper investigates the detection of 3D reflection symmetry from a single RGB image, and reveals its significant benefit on single-image 3D generation. We…

Cited by 0SourcePDFScholar
2024

ZeroShape: Regression-based Zero-shot Shape Reconstruction

CVPR 2024poster

We study the problem of single-image zero-shot 3D shape reconstruction. Recent works learn zero-shot shape reconstruction through generative modeling of 3D assets but these models are computationally expensive at train and inference time. In contrast the traditional approach to this problem is regre…

2023

ShapeClipper: Scalable 3D Shape Learning From Single-View Images via Geometric and CLIP-Based Consistency

CVPR 2023poster

We present ShapeClipper, a novel method that reconstructs 3D object shapes from real-world single-view RGB images. Instead of relying on laborious 3D, multi-view or camera pose annotation, ShapeClipper learns shape reconstruction from a set of single-view segmented images. The key idea is to facilit…

Cited by 22SourcePDFScholar
2022

Planes vs. Chairs: Category-Guided 3D Shape Learning without Any 3D Cues

ECCV 2022poster

"We present a novel 3D shape reconstruction method which learns to predict an implicit 3D shape representation from a single RGB image. Our approach uses a set of single-view images of multiple object categories without viewpoint annotation, forcing the model to learn across multiple object categori…

Cited by 16SourcePDFScholar