← Search

Shijia Kang

3 accepted papers

2025

Beyond Single-Task: Robust Multi-Task Length Generalization for LLMs

NeurIPS 2025poster

Length generalization—the ability to solve problems longer than those seen during training—remains a critical challenge for large language models (LLMs). Previous work modifies positional encodings (PEs) and data formats to improve length generalization on specific symbolic tasks such as addition an…

Cited by 0SourceScholar
2025

Number Cookbook: Number Understanding of Language Models and How to Improve It

ICLR 2025poster

Large language models (LLMs) can solve an increasing number of complex reasoning tasks while making surprising mistakes in basic numerical understanding and processing (such as $9.11 > 9.9$). The latter ability is essential for tackling complex arithmetic and mathematical problems and serves as a fo…

2025

On the Completeness of Invariant Geometric Deep Learning Models

ICLR 2025poster

Invariant models, one important class of geometric deep learning models, are capable of generating meaningful geometric representations by leveraging informative geometric features in point clouds. These models are characterized by their simplicity, good experimental results and computational effici…