← Search

Junbo Guo

6 accepted papers

2025

Leveraging Importance Sampling to Detach Alignment Modules from Large Language Models

NeurIPS 2025poster

The widespread adoption of large language models (LLMs) across industries has increased the demand for high-quality and customizable outputs. However, traditional alignment methods often require retraining large pretrained models, making it difficult to quickly adapt and optimize LLMs for diverse ap…

Cited by 0SourceScholar
2025

On-the-fly Preference Alignment via Principle-Guided Decoding

ICLR 2025poster

With the rapidly expanding landscape of large language models, aligning model generations with human values and preferences is becoming increasingly important. Popular alignment methods, such as Reinforcement Learning from Human Feedback, have shown significant success in guiding models with greater…

2024

LIRE: listwise reward enhancement for preference alignment

ACL 2024findings

Recently, tremendous strides have been made to align the generation of Large Language Models (LLMs) with human values to mitigate toxic or unhelpful content. Leveraging Reinforcement Learning from Human Feedback (RLHF) proves effective and is widely adopted by researchers. However, implementing RLHF…

2022

Improving Chinese Spelling Check by Character Pronunciation Prediction: The Effects of Adaptivity and Granularity

EMNLP 2022main

Chinese spelling check (CSC) is a fundamental NLP task that detects and corrects spelling errors in Chinese texts. As most of these spelling errors are caused by phonetic similarity, effectively modeling the pronunciation of Chinese characters is a key factor for CSC. In this paper, we consider intr…

2019

Near-infrared Image Guided Neural Networks for Color Image Denoising

ICASSP 2019accepted

Noisy color image and guided near-infrared (NIR) image can be jointly employed to eliminate noise and enhance details. Existing methods mostly rely on explicit designed filters and hand-crafted objective function optimization. These methods usually introduce erroneous structures from guidance signal…

Cited by 0SourceScholar