ArXiv will issue a 1-year ban to authors who submit papers containing unchecked LLM output, computer science section chair Thomas Dietterich announced Thursday. The preprint server is treating the policy as a one-strike rule, with a follow-on requirement that any later arXiv submissions from a banned author first pass a reputable peer-reviewed venue. The trigger isn't using a language model — it's leaving the model's mistakes in the manuscript.
Dietterich laid out the standard plainly: "if a submission contains incontrovertible evidence that the authors did not check the results of LLM generation, this means we can't trust anything in the paper." The qualifying evidence includes hallucinated references and stray comments to or from the LLM still embedded in the text. Moderators flag the issue, section chairs confirm the evidence, and authors can appeal before the ban takes effect.
The policy is narrower than an outright ban on AI tools. Dietterich framed it as a demand that authors take "full responsibility" for the content, "irrespective of how the contents are generated." Copy-pasting inappropriate language, plagiarized passages, biased material, errors, or fabricated citations from a model is still the author's problem, not the model's.
Key facts
- 01ArXiv will impose a 1-year ban on authors caught submitting papers with unchecked LLM-generated content.
- 02After the ban, those authors must clear a peer-reviewed venue before posting to arXiv again.
- 03Computer science chair Thomas Dietterich announced the one-strike rule Thursday.
- 04Evidence triggering the rule includes hallucinated references and visible prompts to or from the LLM.
- 05ArXiv is leaving Cornell after more than 20 years of hosting to become an independent nonprofit.
ArXiv has tightened its filters in stages as AI-generated submissions climbed. The site already requires first-time posters to secure an endorsement from an established author before they can submit, a friction layer aimed at the lowest-effort spam. The new ban escalates the consequences for researchers who clear that bar and still ship sloppy work.
“If a submission contains incontrovertible evidence that the authors did not check the results of LLM generation, this means we can't trust anything in the paper.”— Jaeden Schafer
The repository's structural moment matters here too. After more than 20 years hosted by Cornell, arXiv is becoming an independent nonprofit, a shift designed to let it raise money against operational problems including AI-generated noise. Moderation policy is one of the first visible outputs of that transition.
The underlying problem is documented. Recent peer-reviewed research has flagged a rise in fabricated citations in biomedical literature, with language models the most plausible cause. ArXiv sits upstream of formal publication in computer science and math, where preprints often circulate as the primary version of a result for months, so unchecked hallucinations there propagate fast.
This is the second arXiv-related story AI Chat Daily has covered in recent weeks, following coverage of AI-written research papers passing peer review at venues including Scientific Reports. The throughline: editorial systems built for a world of human-paced error are absorbing machine-paced error, and the institutions sitting at the front of the pipeline are the first to write rules.
The enforcement design tries to keep the bar high. Requiring a section chair to confirm flagged evidence, and offering an appeal, suggests arXiv wants to avoid mass auto-bans from pattern-matching on phrases like "as an AI language model." In practice the most catchable cases are the obvious ones — fabricated DOIs, references to papers that don't exist, leftover ChatGPT scaffolding — rather than subtle factual drift inside an otherwise plausible manuscript.
That leaves the harder question untouched. A paper with clean-looking but subtly wrong LLM-generated claims, no hallucinated citations, and no telltale prompt residue will not trip this rule. Dietterich's standard is about visible negligence, not about catching every AI-assisted error, and arXiv has not described any automated detection layer behind the human review.
For AI labs, the policy is a small but pointed data point on where the scientific establishment is drawing lines around model output. ArXiv isn't banning the tools its own community helped build — it's pricing the cost of using them carelessly. Expect more venues to copy the structure, because a one-strike rule with an appeal path is cheaper to defend than a blanket prohibition and easier to enforce than a quality bar.
Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.
Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.



