← Search

Rana Khan

2 accepted papers

2026

Does AI Reviewer See the Full Picture? Attacking and Defending Multimodal Peer Review

ICML 2026poster

The formal integration of Large Language Models (LLMs) and Multimodal LLMs (MLLMs) into scientific peer-review workflows introduces novel and significant risks. Their safety against adversarial manipulation remains critically underexplored, especially given the multimodal nature of scientific papers…

Cited by 0SourceScholar
2026

TMS: Trajectory-Mixed Supervision for Reward-Free, On-Policy SFT

ICML 2026poster

Reinforcement Learning (RL) and Supervised Fine-Tuning (SFT) are the two dominant paradigms for enhancing Large Language Model (LLM) performance on downstream tasks. While RL generally preserves broader model capabilities (retention) better than SFT, it comes with significant costs: complex reward e…

Cited by 0SourceScholar