← Search

Chengwei Dai

4 accepted papers

2025

Capture the Key in Reasoning to Enhance CoT Distillation Generalization

ACL 2025long

As Large Language Models (LLMs) scale up and gain powerful Chain-of-Thoughts (CoTs) reasoning abilities, practical resource constraints drive efforts to distill these capabilities into more compact Smaller Language Models (SLMs). We find that CoTs consist mainly of simple reasoning forms, with a sma…

2024

Improve Student’s Reasoning Generalizability through Cascading Decomposed CoTs Distillation

EMNLP 2024main

Large language models (LLMs) exhibit enhanced reasoning at larger scales, driving efforts to distill these capabilities into smaller models via teacher-student learning.Previous works simply fine-tune student models on teachers’ generated Chain-of-Thoughts (CoTs) data. Although these methods enhance…

2023

CT-GAT: Cross-Task Generative Adversarial Attack based on Transferability

EMNLP 2023long main

Neural network models are vulnerable to adversarial examples, and adversarial transferability further increases the risk of adversarial attacks. Current methods based on transferability often rely on substitute models, which can be impractical and costly in real-world scenarios due to the unavailab…

Cited by 0SourcecodeScholar