← Search

Yuying Xie

3 accepted papers

2025

Cross-Domain Graph Data Scaling: A Showcase with Diffusion Models

NeurIPS 2025poster

Models for natural language and images benefit from data scaling behavior: the more data fed into the model, the better they perform. This 'better with more' phenomenon enables the effectiveness of large-scale pre-training on vast amounts of data. However, current graph pre-training methods struggle…

Cited by 0SourcecodeScholar
2024

CellPLM: Pre-training of Cell Language Model Beyond Single Cells

ICLR 2024poster

The current state-of-the-art single-cell pre-trained models are greatly inspired by the success of large language models. They trained transformers by treating genes as tokens and cells as sentences. However, three fundamental differences between single-cell data and natural language data are overlo…

Cited by 24SourcePDFScholar
2019

Manifold denoising by Nonlinear Robust Principal Component Analysis

NeurIPS 2019poster

This paper extends robust principal component analysis (RPCA) to nonlinear manifolds. Suppose that the observed data matrix is the sum of a sparse component and a component drawn from some low dimensional manifold. Is it possible to separate them by using similar ideas as RPCA? Is there any benefit…

Cited by 19SourcePDFScholar