Study traces LLM hallucinations to competing latent associations
A new arXiv paper examines why large language models produce incorrect answers even when the correct knowledge is present in their training data. The authors frame inference as competition between latent associations, where the wrong "key" can win out during answer selection, and use this view to identify and flag hallucinated outputs.