papersTODAY 04:00 UTC
Study finds biomedical reference generation unreliable across 26 LLMs
Researchers tested 26 large language models from eight developers on their ability to produce accurate biomedical citations. The authors report that fabricated or incorrect references remain a persistent problem across the models evaluated. The work is a preprint and characterises the frequency of this failure mode rather than proposing a fix.