AI models often give the right answers but point to the wrong sources
Source published ·Modelwire updated
Original coverage: The Decoder ↗·How Modelwire adds context

The development
A systematic gap has emerged in how leading language models justify their outputs. Researchers at Peking University documented that GPT and Gemini frequently cite document passages that don't actually support their conclusions, even when final answers prove correct. This 'attribution hallucination' poses material risk in regulated domains like law and medicine where reasoning transparency is non-negotiable. The new CiteVQA benchmark provides the first standardized test for this failure mode, shifting evaluation focus from answer accuracy alone to the integrity of supporting evidence chains.
Modelwire’s AI-generated summary of coverage from The Decoder.
Modelwire analysis
ExplainerOur AI-generated reading of the wider context and the next developments to watch.
The more unsettling implication buried in this finding is that attribution hallucination is nearly invisible in standard evaluations: models score well on answer accuracy while silently fabricating their evidentiary basis, meaning current deployment safeguards in high-stakes domains may be measuring the wrong thing entirely.
This is largely disconnected from recent activity in our archive, as Modelwire has no prior coverage to anchor it to. It belongs, however, to a broader and well-documented conversation in the research community about the gap between benchmark performance and real-world reliability. The specific failure mode here sits adjacent to hallucination research but is meaningfully different: the model is not wrong about the world, it is wrong about its own reasoning trail. That distinction matters most in legal and medical contexts, where a correct conclusion built on a fabricated citation chain can expose practitioners to liability even when the underlying answer holds up.
Watch whether legal-tech or clinical AI vendors that currently cite GPT or Gemini in their compliance materials respond to CiteVQA with their own attribution audits within the next two quarters. Silence from that segment would itself be informative.
This interpretation is generated from the summary above and available source metadata. Our methodology · Report an error
MentionsGPT · Gemini · Peking University · CiteVQA
How this coverage is produced
Modelwire uses AI to generate summaries and context from source headlines, snippets, and selected archive coverage. Automated checks do not verify every claim, and items are not routinely reviewed by a person before publication. Zacaria Solis operates the site. Read the linked source for the full evidence and report errors through our corrections process.
Modelwire summarizes, we don’t republish. The full content lives on the-decoder.com. If you’re a publisher and want a different summarization policy for your work, see our takedown page.