Signal2026-07-21
arXiv

How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs?

Part of

Context Sensitivity In Generative AI