← Search

Zhiyuan Ren

5 accepted papers

2024

Distilling CLIP with Dual Guidance for Learning Discriminative Human Body Shape Representation

CVPR 2024poster

Person Re-Identification (ReID) holds critical importance in computer vision with pivotal applications in public safety and crime prevention. Traditional ReID methods reliant on appearance attributes such as clothing and color encounter limitations in long-term scenarios and dynamic environments. To…

Cited by 12SourcePDFScholar
2024

TIGER: Time-Varying Denoising Model for 3D Point Cloud Generation with Diffusion Process

CVPR 2024poster

Recently diffusion models have emerged as a new powerful generative method for 3D point cloud generation tasks. However few works study the effect of the architecture of the diffusion model in the 3D point cloud resorting to the typical UNet model developed for 2D images. Inspired by the wide adopti…

2023

ChatGPT-Powered Hierarchical Comparisons for Image Classification

NeurIPS 2023poster

The zero-shot open-vocabulary setting poses challenges for image classification. Fortunately, utilizing a vision-language model like CLIP, pre-trained on image-text pairs, allows for classifying images by comparing embeddings. Leveraging large language models (LLMs) such as ChatGPT can further enhan…

2023

Diffusion Motion: Generate Text-Guided 3D Human Motion by Diffusion Model

ICASSP 2023accepted

We propose a simple and novel method for generating 3D human motion from complex natural language sentences, which describe different velocity, direction and composition of all kinds of actions. Different from existing methods that use classical generative architecture, we apply the Denoising Diffus…

Cited by 0SourceScholar
2023

Hierarchical Fine-Grained Image Forgery Detection and Localization

CVPR 2023poster

Differences in forgery attributes of images generated in CNN-synthesized and image-editing domains are large, and such differences make a unified image forgery detection and localization (IFDL) challenging. To this end, we present a hierarchical fine-grained formulation for IFDL representation learn…