← Search

Qing Song

4 accepted papers

2025

Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation

CVPR 2025poster

In-domain generation aims to perform a variety of tasks within a specific domain, such as unconditional generation, text-to-image, image editing, 3D generation, and more. Early research typically required training specialized generators for each unique task and domain, often relying on fully-labeled…

Cited by 0SourcePDFScholar
2023

Large-Scale Person Detection and Localization Using Overhead Fisheye Cameras

ICCV 2023oral

Location determination finds wide applications in daily life. Instead of existing efforts devoted to localizing tourist photos captured by perspective cameras, in this article, we focus on developing person positioning solutions using overhead fisheye cameras. Such solutions are advantageous in larg…

Cited by 25PDFScholar
2020

Renovating Parsing R-CNN for Accurate Multiple Human Parsing

ECCV 2020poster

Multiple human parsing aims to segment various human parts and associate each part with the corresponding instance simultaneously. This is a very challenging task due to the diverse human appearance, semantic ambiguity of different body parts and clothing, and complex background. Through analysis of…