← Search

Jian-Tao Huang

2 accepted papers

2025

The Emperor’s New Reasoning: Format Imitation Overshadows Genuine Mathematical Understanding in SFT

EMNLP 2025

Recent advances in large language models (LLMs) have yielded impressive gains on mathematical reasoning benchmarks via supervised fine-tuning (SFT). However, the brittleness of these models under input perturbations has cast doubt on whether such improvements reflect genuine reasoning abilities or m

Cited by 0SourcePDFScholar
2024

NumHG: A Dataset for Number-Focused Headline Generation

COLING 2024main

Headline generation, a key task in abstractive summarization, strives to condense a full-length article into a succinct, single line of text. Notably, while contemporary encoder-decoder models excel based on the ROUGE metric, they often falter when it comes to the precise generation of numerals in h…