← Search

Bokai Hu

1 accepted papers

2025

Improving the Language Understanding Capabilities of Large Language Models Using Reinforcement Learning

EMNLP 2025

Instruction-fine-tuned large language models (LLMs) under 14B parameters continue to underperform on natural language understanding (NLU) tasks, often trailing smaller models like BERT-base on benchmarks such as GLUE and SuperGLUE. Motivated by the success of reinforcement learning in reasoning task