← Search

Asher Parker-Sartori

2 accepted papers

2025

Diverse Preference Learning for Capabilities and Alignment

ICLR 2025poster

As LLMs increasingly impact society, their ability to represent diverse perspectives is critical. However, recent studies reveal that alignment algorithms such as RLHF and DPO significantly reduce the diversity of LLM outputs. Not only do aligned LLMs generate text with repetitive structure and wor…

Cited by 1SourcePDFScholar
2025

On the creation of narrow AI: hierarchy and nonlocality of neural network skills

NeurIPS 2025poster

We study the problem of creating strong, yet narrow, AI systems. While recent AI progress has been driven by the training of large general-purpose foundation models, the creation of smaller models specialized for narrow domains could be valuable for both efficiency and safety. In this work, we explo…

Cited by 0SourcecodeScholar