
ModelsLead story
LLMs believe false claims even when training data labels them as lies
A new preprint finds Qwen, Kimi, and GPT-4.1 absorb fabricated facts at an 88.6% belief rate even after explicit negation warnings.
Jaeden SchaferEditor in Chief5 min read