Turning Bias into Bugs: Bandit-Guided Style Manipulation Attacks on LLM Judges
Large Language Models (LLMs) are increasingly employed as automated judges for evaluating generative models. However, their known stylistic biases, such as a preference for verbosity or specific sentence structures, present an underexplored security vulnerability. In this work, we introduce BITE (BI…