← Search

Robert Krzyzanowski

1 accepted papers

2026

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

ICML 2026poster

Recent studies indicate that when faced with explicit biases in prompts, models often omit mentioning these biases in their Chain-of-Thought (CoT) output, revealing that verbalized reasoning can give an incorrect picture of how models arrive at conclusions (unfaithfulness). In this work, we show tha…

Cited by 0SourcecodeScholar