← Search

Jilong Gao

1 accepted papers

2025

On-Policy Optimization with Group Equivalent Preference for Multi-Programming Language Understanding

NeurIPS 2025poster

Large language models (LLMs) achieve remarkable performance in code generation tasks. However, a significant performance disparity persists between popular programming languages (e.g., Python, C++) and others. To address this capability gap, we leverage the code translation task to train LLMs, ther…

Cited by 0SourceScholar