← Search

Eric Gilbert

2 accepted papers

2026

Position: AI/ML Deepfake Research is Misaligned with AI Generated Non-Consensual Intimate Imagery (AIG-NCII)

ICML 2026oral

AI-generated non-consensual intimate imagery (AIG-NCII) is not adequately addressed in AI/ML literature regarding AI-generated media, commonly referred to as "deepfakes". While research on deepfakes currently focuses on its epistemic harms—or harms relating to truth and authenticity—this is misalign…

Cited by 0SourceScholar
2025

Deep Value Benchmark: Measuring Whether Models Generalize Deep values or Shallow Preferences

NeurIPS 2025spotlight

We introduce the Deep Value Benchmark (DVB), an evaluation framework that directly tests whether large language models (LLMs) learn fundamental human values or merely surface-level preferences. This distinction is critical for AI alignment: Systems that capture deeper values are likely to generalize…

Cited by 0SourceScholar