← Search

Vinit Ravishankar

4 accepted papers

2022

A Closer Look at Parameter Contributions When Training Neural Language and Translation Models

COLING 2022main

We analyze the learning dynamics of neural language and translation models using Loss Change Allocation (LCA), an indicator that enables a fine-grained analysis of parameter updates when optimizing for the loss function. In other words, we can observe the contributions of different network component…

2022

Word Order Does Matter and Shuffled Language Models Know It

ACL 2022long

Recent studies have shown that language models pretrained and/or fine-tuned on randomly permuted sentences exhibit competitive performance on GLUE, putting into question the importance of word order information. Somewhat counter-intuitively, some of these studies also report that position embeddings…

Cited by 49SourcePDFScholar