← Search

Zixuan Xie

5 accepted papers

2026

MathlibLemma: Folklore Lemma Generation and Benchmark for Formal Mathematics

ICML 2026poster

While the ecosystem of Lean and Mathlib has enjoyed celebrated success in formal mathematical reasoning with the help of large language models (LLMs), the absence of many folklore lemmas in Mathlib remains a persistent barrier that limits Lean's usability as an everyday tool for mathematicians like …

Cited by 0SourceScholar
2026

Offline Two-Player Zero-Sum Markov Games with KL Regularization

ICML 2026poster

We study the problem of learning Nash equilibria in offline two-player zero-sum Markov games. While existing approaches often rely on explicit pessimism to address distribution shift, we show that KL regularization alone suffices to stabilize learning and guarantee convergence. We first introduce Re…

Cited by 0SourceScholar
2025

Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features

NeurIPS 2025poster

Linear TD($\lambda$) is one of the most fundamental reinforcement learning algorithms for policy evaluation. Previously, convergence rates are typically established under the assumption of linearly independent features, which does not hold in many practical scenarios. This paper instead establishes…

Cited by 0SourceScholar
2025

IntrinsicControlNet: Cross-distribution Image Generation with Real and Unreal

ICCV 2025poster

Realistic images are usually produced by simulating light transportation results of 3D scenes using rendering engines. This framework can precisely control the output but is usually weak at producing photo-like images. Alternatively, diffusion models have seen great success in photorealistic image g…

Cited by 0SourcePDFScholar
2025

Linear $Q$-Learning Does Not Diverge in $L^2$: Convergence Rates to a Bounded Set

ICML 2025poster

$Q$-learning is one of the most fundamental reinforcement learning algorithms. It is widely believed that $Q$-learning with linear function approximation (i.e., linear $Q$-learning) suffers from possible divergence until the recent work Meyn (2024) which establishes the ultimate almost sure boundedn…

Cited by 0SourcePDFScholar