2026
Get RICH or Die Scaling: Profitably Trading Inference Compute for Robustness
ICLR 2026poster
Models are susceptible to adversarially out-of-distribution (OOD) data despite large training-compute investments into their robustification. Zaremba et al. (2025) make progress on this problem at test time, showing LLM reasoning improves satisfaction of model specifications designed to thwart attac…