ArXiv Bans Authors 1 Year for Unchecked AI Output — Fake Citations Up 10x Since 2023
ArXiv has formalised what was always technically true: authors are responsible for everything in their papers, including text their LLM generated. The difference now is a concrete penalty — a one-year ban from submitting to the platform, followed by a permanent requirement to clear peer review before any future preprint posts.
Thomas Dietterich, chair of arXiv’s computer science section, announced the policy on May 14 in a thread addressed directly to authors. The trigger is narrow and deliberate: “incontrovertible evidence” that an author did not check the LLM’s output before submitting. Hallucinated references that point to no real paper. Meta-comments from the model left in the manuscript body — phrases like “here is a 200-word summary; would you like me to make any changes?” — are the paradigmatic example. Routine AI-assisted editing is not targeted. The rule catches papers whose authors demonstrably did not read them.
Enforcement requires a two-step confirmation before any ban sticks: a moderator must first flag the evidence, and a section chair must confirm it. There is an appeal process. But once confirmed, the consequence is the platform’s most serious sanction.
Why It Matters Now
The scale of the problem drove this. A Lancet study published in May 2026 audited 2.5 million biomedical papers and 126 million references indexed on PubMed Central. Fabricated citations in 2023 appeared at a rate of roughly one paper in 2,828. By early 2026, the rate had reached one in 277 — a tenfold increase in three years. The researchers tied the acceleration directly to the proliferation of AI writing tools.
At NeurIPS 2025, a GPTZero scan of 4,841 accepted papers — papers that had cleared multi-reviewer evaluation — found more than 100 hallucinated citations across 53 papers. The reviewers did not catch them. This is the practical consequence of hallucinated citations being difficult to detect without manually verifying every reference.
ArXiv is not a journal. It does not peer-review its papers. But for machine learning and astrophysics, it is the de facto distribution channel for research. Papers posted to arXiv are read, cited, and built upon before formal peer review, which means a hallucinated citation can propagate through the literature just as fast as a real one, and often faster.
What Counts as Evidence
Dietterich listed three archetypal cases:
- Hallucinated references — citations that correspond to no real publication
- Model meta-comments — chatbot dialogue or instructions from the LLM left in the manuscript body
- Placeholder data — tables or figures with notes telling the author to “fill in with real numbers” that were never removed
All three are identifiable without AI-detection tools, which the policy deliberately avoids relying on. Detecting whether a well-edited paper was AI-drafted is technically unreliable. Detecting a model’s offer to revise its own output left in the body text is not.
Context and What Comes Next
ArXiv is simultaneously undergoing a structural transition, moving from a Cornell-hosted project to an independent nonprofit after more than 30 years. That shift should give it more operational autonomy and the ability to fund enforcement infrastructure directly.
The platform already introduced an endorsement requirement for first-time submitters in 2025, requiring a voucher from an established researcher before first-time posting is permitted. The new ban extends accountability into a per-author sanction across the platform rather than just a gatekeeping step at entry.
The first documented one-year ban under the new threshold is the next concrete gate. Until it is imposed and publicly acknowledged, the policy’s deterrent effect is theoretical. Scientific Director Steinn Sigurðsson has described what arXiv rejects: “some of it is really really egregious.” The question now is whether a formal penalty changes the calculus for researchers who have been treating arXiv as a consequence-free dumping ground.
Key Numbers
- Fake citation rate (2023): 1 in 2,828 papers
- Fake citation rate (early 2026): 1 in 277 papers — 10x increase
- NeurIPS 2025: 100+ hallucinated citations found in 53 accepted, peer-reviewed papers
- Lancet study scope: 2.5 million papers, 126 million references
- Penalty: 1-year arXiv submission ban, then mandatory peer review for all future submissions
- Trigger: Incontrovertible evidence of unchecked LLM output only — not AI-assisted editing