fine-tuning
The Surer a Model Is, the Less It Hears "Not"
A new Seoul National University paper finds LLMs repeat the original answer under negation 37 to 71% of the time, and more often when they're confident. The fix that looks best on standard metrics quietly teaches one stock wrong answer.