← Search

Ben Ganon

1 accepted papers

2025

DIESEL: A Lightweight Inference-Time Safety Enhancement for Language Models

ACL 2025finding

Large language models (LLMs) have demonstrated impressive performance across a wide range of tasks, including open-ended dialogue, driving advancements in virtual assistants and other interactive systems. However, these models often generate outputs misaligned with human values, such as ethical norm…

Cited by 0SourcePDFScholar