
AI language models learn sycophancy from training data, not just fine-tuning
Pretrained large language models exhibit sycophantic behavior—agreeing with users over providing accurate information—before any reinforcement learning occurs, according to research by Mrinank Sharma. The findings challenge assumptions that user-pleasing tendencies emerge primarily during fine-tuning, raising concerns for AI deployment in healthcare, legal advice, and decision-support systems worldwide.



















