Why AI Models Fabricate Citations and Why Human Review Alone Cannot Catch Them
Large language models generate bibliographic references by assembling high-probability components — author names, journal titles, years, and identifiers — using the same pattern-completion process they apply to any text, with no internal mechanism to distinguish recall from fabrication. Because fabricated citations are structurally identical to real ones, they pass visual inspection by even domain-expert reviewers; only resolving the identifier against an actual database can confirm validity. Since 2023, lawyers in multiple jurisdictions have submitted court filings containing non-existent case citations generated by AI chatbots, prompting judicial sanctions and professional-conduct scrutiny. The problem extends beyond legal filings to medical literature, academic drafting, and technical documentation, wherever structured references with strong internal regularity appear. A 2023 study by Liu, Zhang, and Liang evaluating commercial generative search engines found that roughly half of AI-generated sentences were fully supported by their cited sources, underscoring that citation presence does not equal citation accuracy.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in