← Search

Ran Zhang

17 accepted papers

2026

D&R: Recovery-based AI-Generated Text Detection via a Single Black-box LLM Call

ICLR 2026poster

Large language models (LLMs) generate increasingly human-like text, raising concerns about misinformation and authenticity. Detecting AI-generated text remains challenging: existing methods often underperform, especially on short texts, require probability access unavailable in real-world black-box…

Cited by 0SourcecodeScholar
2026

Discriminative Mixture-of-Experts on Graphs with Reliable Expert Fusion

ICML 2026poster

Graph Mixture-of-Experts (Graph-MoE) offers a way to scale GNNs via adaptive capacity allocation, with the goal of allowing different experts to capture diverse graph patterns. Its effectiveness heavily depends on the coordination between routing decisions and expert specialization. However, through…

Cited by 0SourceScholar
2026

Invariant Conditional Molecular Generation Under Distribution Shift

AAAI 2026technical

Conditional molecular generation, aiming to generate 2D and 3D molecules that satisfy given properties, has achieved remarkable progress, thanks to the advances in deep generative models such as graph diffusion. However, existing methods generally assume that the given conditions for training and te

Cited by 0SourcePDFScholar
2026

scCluBench: Comprehensive Benchmarking of Clustering Algorithms for Single-Cell RNA Sequencing

AAAI 2026technical

Cell clustering is crucial for uncovering cellular heterogeneity in single-cell RNA sequencing (scRNA-seq) data by identifying cell types and marker genes. Despite its importance, existing benchmarks for scRNA-seq clustering remain fragmented, lacking standardized protocols and often omitting recent

Cited by 0SourcePDFScholar
2025

GeoRVLF: A Robust Drone-Satellite Visual Geo-Localization Framework for Small Unmanned Aerial Vehicle Platforms

RA-L 2025

Drone-satellite geo-localization is a novel technology for autonomous UAV positioning in GNSS-denied environments, but it faces many complex challenges such as cross-scale heterogeneous scenes and real-time operational efficiency in practical applications. This letter proposes a robust visual geoloc

Cited by 1SourceScholar
2025

How Good Are LLMs for Literary Translation, Really? Literary Translation Evaluation with Humans and LLMs

NAACL 2025long

Recent research has focused on literary machine translation (MT) as a new challenge in MT. However, the evaluation of literary MT remains an open problem. We contribute to this ongoing discussion by introducing LITEVAL-CORPUS, a paragraph-level parallel corpus containing verified human translations…

2025

LiTransProQA: An LLM-based Literary Translation Evaluation Metric with Professional Question Answering

EMNLP 2025

The impact of Large Language Models (LLMs) has extended into literary domains. However, existing evaluation metrics for literature prioritize mechanical accuracy over artistic expression and tend to overrate machine translation as being superior to human translation from experienced professionals. I

2025

Make-It-Animatable: An Efficient Framework for Authoring Animation-Ready 3D Characters

CVPR 2025highlight

3D characters are essential to modern creative industries, but making them animatable often demands extensive manual work in tasks like rigging and skinning. Existing automatic rigging tools face several limitations, including the necessity for manual annotations, rigid skeleton topologies, and limi…

2025

Motif-Oriented Representation Learning with Topology Refinement for Drug-Drug Interaction Prediction

AAAI 2025technical

Drug-Drug Interaction (DDI) prediction has attracted considerable attention in designing multi-drug combination strategies and avoiding adverse reactions. Notably, Artificial Intelligence (AI)-driven DDI prediction methods have emerged as a pivotal research paradigm. However, most AI-driven DDI pred…

Cited by 0SourcePDFScholar
2025

Time-Aware Auto White Balance in Mobile Photography

ICCV 2025poster

Cameras rely on auto white balance (AWB) to correct undesirable color casts caused by scene illumination and the camera's spectral sensitivity. This is typically achieved using an illuminant estimator that determines the global color cast solely from the color information in the camera's raw sensor…

Cited by 0SourcePDFScholar
2024

PolitiCause: An Annotation Scheme and Corpus for Causality in Political Texts

COLING 2024main

In this paper, we present PolitiCAUSE, a new corpus of political texts annotated for causality. We provide a detailed and robust annotation scheme for annotating two types of information: (1) whether a sentence contains a causal relation or not, and (2) the spans of text that correspond to the cause…

2024

Robust Decoding of the Auditory Attention from EEG Recordings Through Graph Convolutional Networks

ICASSP 2024accepted

Auditory attention decoding (AAD) with electroencephalography (EEG) holds great promise in brain-computer interface (BCI). Despite much progress, it remains a research topic on how to effectively evaluate the performance of EEG-based AAD algorithms under an appropriate setting that reflects the use…

Cited by 0SourceScholar
2023

Into the Single Cell Multiverse: an End-to-End Dataset for Procedural Knowledge Extraction in Biomedical Texts

NeurIPS 2023spotlight

Many of the most commonly explored natural language processing (NLP) information extraction tasks can be thought of as evaluations of declarative knowledge, or fact-based information extraction. Procedural knowledge extraction, i.e., breaking down a described process into a series of steps, has rece…

Cited by 1SourcePDFScholar
2023

StepFormer: Self-Supervised Step Discovery and Localization in Instructional Videos

CVPR 2023poster

Instructional videos are an important resource to learn procedural tasks from human demonstrations. However, the instruction steps in such videos are typically short and sparse, with most of the video being irrelevant to the procedure. This motivates the need to temporally localize the instruction s…

Cited by 29SourcePDFScholar
2022

IKEA-Manual: Seeing Shape Assembly Step by Step

NeurIPS 2022accept

Human-designed visual manuals are crucial components in shape assembly activities. They provide step-by-step guidance on how we should move and connect different parts in a convenient and physically-realizable way. While there has been an ongoing effort in building agents that perform assembly tasks…

Cited by 19SourcePDFScholar
2022

Melons: Generating Melody With Long-Term Structure Using Transformers And Structure Graph

ICASSP 2022accepted

The creation of long melody sequences requires effective expression of coherent musical structure. However, there is no clear representation of musical structure. Recent works on music generation have suggested various approaches to deal with the structural information of music, but generating a ful…

Cited by 0SourceScholar