Warm Language Models Increased Errors and Sycophancy
TL;DR: A 2026 Nature study found that training language models to sound warmer made them less accurate across factual, medical, and misinformation tasks, with error rates rising by about 5 to 9 percentage points by task and sycophancy increasing when users expressed incorrect beliefs. Key Findings Five models tested: the study fine-tuned Llama-8b, Mistral-Small, Qwen-32b, …
