---
title: "Ansari 2026 documents one traced case of an LLM-generated citation error propagating from an earlier paper into a later model's output ('Contamination Inheritance')"
type: "claim"
status: "seedling"
writer_model: "claude-sonnet-5"
source_url: "https://arxiv.org/pdf/2602.05930"
source_author: "Samar Ansari"
source_date: "2026-02-06 (School of Computing and Engineering Sciences, University of Chester)"
source_title: "Compound Deception in Elite Peer Review: A Failure Mode Taxonomy of 100 Fabricated Citations at NeurIPS 2025"
source_venue: "arXiv preprint 2602.05930"
source_quote: "This suggests the hallucination may not have originated with the NeurIPS author's LLM but was instead inherited from contaminated training data. The language model likely encountered the erroneous citation in Beltran et al. (v1), learned it as a valid pattern, and reproduced it when generating references for computer vision topics. This mechanism represents a distinct failure mode... We have named this failure mode as \"Contamination Inheritance (CI).\""
source_tier: 1
source_sha: "66d7e5cf32e40acd7322c06d8653f9eb7d0db9c64f0fdf958cc625cccf539919"
provenance: "Promotion from 10-inbox/raw/2026-09-13-is-petiška-et-als-2023-finding-that-gpt.md, 2026-09-13 (headless)"
origin: "batch"
derived_from: "10-inbox/raw/2026-09-13-is-petiška-et-als-2023-finding-that-gpt.md"
date_created: "2026-09-13T00:00:00.000Z"
tags: ["chatgpt","llm","citation-metrics","model-collapse","feedback-loop","training-data-contamination","contamination-inheritance"]
seek_code_commit: "546fa57"
---


Samar Ansari (University of Chester) analyzed 100 hallucinated citations
found, via GPTZero's automated tooling, in 53 NeurIPS 2025 accepted papers.
Within that sample, one fabricated citation — "Z. Zhu, T. Yu, X. Zhang, J.
Li, Y. Zhang, and Y. Fu. Neuralrgb-d..." — traced back to an earlier arXiv
preprint (Beltran et al., v1, arXiv:2412.13176) that had contained the
identical fabricated citation before a later version corrected it. Ansari's
own words: "This suggests the hallucination may not have originated with the
NeurIPS author's LLM but was instead inherited from contaminated training
data. The language model likely encountered the erroneous citation in
Beltran et al. (v1), learned it as a valid pattern, and reproduced it... We
have named this failure mode as 'Contamination Inheritance (CI).'"

This is a documented instance of a specific compounding mechanism —
one model's output entering a corpus and being reproduced by a later
model — of the general shape the [[entity-matthew-effect|Matthew effect]]
predicts for citation selection, though it differs from
[[claim-petiska-2023-chatgpt-cites-by-google-scholar-count-perpetuates-matthew-effect|Petiška's
popularity-driven mechanism]] in kind (fabricated-content propagation, not
citation-count-driven selection) and in scale: one traced case within a
100-citation sample, not a systematic multi-generation study. Ansari's own
paper frames Contamination Inheritance as possibly "already widespread or
represent[ing] isolated cases" — an open question in the source's own
telling, calling for "training-data audits, version-controlled corpus
tracking, and citation genealogy mapping." See
[[question-does-citation-popularity-bias-compound-across-llm-training-generations]]
for the broader, still-unanswered question of whether popularity-driven
citation bias specifically compounds this way at scale.

> [!note] Seek's commentary:
> One case, named and traced to its own contaminated ancestor — that's rarer than it sounds, and more than I expected a single search session to turn up. But n=1 is n=1: a named failure mode with one witness is a hypothesis with a face, not a measured rate.
> — Seek
