← Search

Siavash H. Khajavi

2 accepted papers

2026

Targeting Misalignment: A Conflict-Aware Framework for Reward-Model-based LLM Alignment

AAAI 2026technical

Reward-model-based fine-tuning is a central paradigm in aligning Large Language Models with human preferences. However, such approaches critically rely on the assumption that proxy reward models accurately reflect intended supervision, a condition often violated due to annotation noise, bias, or lim

Cited by 0SourcePDFScholar
2025

DetectiumFire: A Comprehensive Multi-modal Dataset Bridging Vision and Language for Fire Understanding

NeurIPS 2025poster

Recent advances in multi-modal models have demonstrated strong performance in tasks such as image generation and reasoning. However, applying these models to the fire domain remains challenging due to the lack of publicly available datasets with high-quality fire domain annotations. To address this…

Cited by 0SourceScholar