← Search

Yun Liao

6 accepted papers

2026

Compositional Generalization from Learned Skills via CoT Training: A Theoretical and Structural Analysis for Reasoning

ICLR 2026poster

Chain-of-Thought (CoT) training has markedly advanced the reasoning capabilities of large language models (LLMs), yet the mechanisms by which CoT training enhances generalization remain inadequately understood. In this work, we demonstrate that compositional generalization is fundamental: models sys…

Cited by 0SourcecodeScholar
2026

Multimodal Gaussian Mixture Variational Autoencoder with Consistency Regularizations

AAAI 2026technical

Variational autoencoder (VAE)-based frameworks possess a natural advantage in modeling the shared and private information inherent in multimodal data. However, current models focus on improving the quality of shared representations from the reconstruction perspective, lacking explicit mechanisms to

Cited by 0SourcePDFScholar
2025

Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning

ICML 2025poster

Test-time scaling, which is also often referred to as *slow-thinking*, has been demonstrated to enhance multi-step reasoning in large language models (LLMs). However, despite its widespread utilization, the mechanisms underlying slow-thinking methods remain poorly understood. This paper explores the…

2025

SDPGO: Efficient Self-Distillation Training Meets Proximal Gradient Optimization

NeurIPS 2025poster

Self-knowledge distillation (SKD) enables single-model training by distilling knowledge from the model's own output, eliminating the need for a separate teacher network required in conventional distillation methods. However, current SKD methods focus mainly on replicating common features in the stud…

Cited by 0SourcecodeScholar
2024

Ahpatron: A New Budgeted Online Kernel Learning Machine with Tighter Mistake Bound

AAAI 2024technical

In this paper, we study the mistake bound of online kernel learning on a budget. We propose a new budgeted online kernel learning model, called Ahpatron, which significantly improves the mistake bound of previous work and resolves an open problem related to upper bounds of hypothesis space constrain…

2018

Streaming Influence Maximization in Social Networks Based on Multi-Action Credit Distribution

ICASSP 2018accepted

In a social network, influence maximization is the problem of identifying a set of users that own the maximum influence ability across the network. In this paper, a novel credit distribution (CD) based model, termed as the multi-action CD (mCD) model, is introduced to quantify the influence ability…

Cited by 0SourceScholar