2025
Token-Level Accept or Reject: A Micro Alignment Approach for Large Language Models
IJCAI 2025
With the rapid development of Large Language Models (LLMs), aligning these models with human preferences and values is critical to ensuring ethical and safe applications. However, existing alignment techniques such as RLHF or DPO often require direct fine-tuning on LLMs with billions of parameters,