Skip to main content
Live
Main content

ArXiv will ban authors for one year over unchecked AI-generated paper content

The preprint server's computer science section chair, Thomas Dietterich, says hallucinated references and leftover LLM meta-comments now trigger a 12-month suspension.

Jaeden Schafer
Editor in Chief · · 4 min read
ArXiv will ban authors for one year over unchecked AI-generated paper content

ArXiv will ban researchers for one year if their submitted papers contain what the preprint server calls incontrovertible evidence that authors did not check the output of large language models. Thomas Dietterich, ArXiv's section chair for computer science, announced the policy on X on May 15, 2026, citing hallucinated references and leftover LLM meta-comments as the kind of evidence that triggers the penalty. After the ban expires, authors will have to get any subsequent ArXiv submission accepted at a reputable peer-reviewed venue before they can post again.

The penalty applies across the full content of a paper. ArXiv's Code of Conduct already holds each named author responsible for everything in a submission, irrespective of how the text was produced. Dietterich's update clarifies the consequence: a single unambiguous trace of unverified model output is enough to invalidate the entire paper in ArXiv's eyes.

Dietterich gave concrete examples of what qualifies as incontrovertible. Hallucinated citations are one. Meta-comments from the model are another, including lines like "here is a 200 word summary; would you like me to make any changes?" or "the data in this table is illustrative, fill it in with the real numbers from your experiments." Both have shown up in submitted manuscripts, indicating the author copy-pasted directly from a chat session without reading what they pasted.

Key facts

  • 01ArXiv will issue a 1-year ban to authors whose papers show incontrovertible evidence of unchecked LLM output, section chair Thomas Dietterich announced May 15, 2026.
  • 02After the ban, authors must get future ArXiv submissions accepted at a peer-reviewed venue before posting.
  • 03Triggering evidence includes hallucinated references and leftover LLM meta-comments such as 'here is a 200 word summary; would you like me to make any changes?'
  • 04Last year ArXiv already restricted computer science review articles and position papers to those accepted at a conference or journal.
  • 05Bans require a moderator to document the issue and the section chair to confirm; authors can appeal.

The internal process has two gates. A moderator first documents the problem, then the section chair confirms it before the ban is imposed. Dietterich told 404Media that authors can appeal a ban decision, and that the policy applies only to cases meeting the incontrovertible-evidence bar — not to papers that merely show signs of AI-assisted writing.

If a submission contains incontrovertible evidence that the authors did not check the results of LLM generation, this means we can't trust anything in the paper. The penalty is a 1-year ban from arXiv.
Jaeden Schafer

ArXiv tightened its rules once already in this direction. Last year, the site updated its policies so that computer science review articles and position papers would only be accepted if they had already been peer-reviewed and accepted at a conference or journal. ArXiv said at the time that large language models had made such content "relatively easy to churn out on demand," and that the majority of review articles it received amounted to "little more than annotated bibliographies, with no substantial discussion of open research issues."

Quoting Dietterich directly: "Our Code of Conduct states that by signing your name as an author of a paper, each author takes full responsibility for all its contents, irrespective of how the contents were generated. If generative AI tools generate inappropriate language, plagiarized content, biased content, errors, mistakes, incorrect references, or misleading content, and that output is included in scientific works, it is the responsibility of the author(s)."

The volume problem is what's forcing the policy. ArXiv operates as a preprint server without traditional peer review, which makes it both the fastest publication channel in machine learning and the easiest to flood. AI Chat Daily covered the parallel problem last week, when AI-written research papers were shown to be passing peer review at established journals — editors and reviewers are losing the ability to filter machine-generated text at the gate.

The one-year ban with a peer-review reentry requirement is a meaningful escalation for a community that treats ArXiv preprints as the primary unit of progress. For computer science and machine learning researchers in particular, a year off ArXiv effectively means a year of invisibility, since conference submissions and citation counts both lean heavily on preprint availability. Forcing reentry through peer review adds further delay on top of that.

Related · from this week
ArXiv will ban authors for a year if LLMs do the work unchecked
Jaeden Schafer · 4 min read →

The harder question is whether moderation can scale. ArXiv receives a large daily submission volume across physics, math, computer science, and related fields, and the incontrovertible-evidence standard puts a human in the loop for every penalty decision. Dietterich's framing — moderator documents, section chair confirms, author can appeal — suggests ArXiv is comfortable applying the rule narrowly and slowly rather than automating detection.

There is also no automated LLM-detection step described in the policy. ArXiv is relying on the giveaways that careless authors leave behind, not on classifier-based screening, which has a poor track record on machine-generated text. That choice avoids false-positive bans of legitimate AI-assisted work, but it also means the enforcement bar lands on the laziest offenders rather than the most prolific ones.

For the broader AI research ecosystem, ArXiv's move sets the first concrete penalty schedule from a major scientific infrastructure provider. Journals have updated authorship policies and conferences have asked authors to disclose AI use, but a one-year posting ban with a peer-review reentry condition is a sharper instrument than either. The signal to authors is straightforward: AI assistance is permitted, but the author signs for the output, and shipping a paper with the model's chat-window scaffolding still attached is now a career-cost event rather than an embarrassment.

ShareXLinkedInEmail
AI Box

Every AI model. One chat.

The latest models from ChatGPT, Claude, Gemini, Sora, ElevenLabs — 80+ models in a single chat. Compare answers side by side. Pick the best one every time.

  • ChatGPT, Claude, Gemini, Grok, DeepSeek — in one chat
  • Generate images & video with Sora, Veo, Ideogram
  • Compare any two models side by side
  • From $8.99/mo · 80+ models, all included
Try AI Boxaibox.ai
Trusted by 3,000+ teams
Got a tip?

Working on something we should cover, or seeing a story we missed? Send leads, documents, or feedback to hello@aichatdaily.com. For sensitive tips, see our secure tips page for Signal and PGP options.

Spotted an error? Email hello@aichatdaily.com with the URL and the issue, or read our full corrections policy.

AI Box Daily briefingFree · Daily · No fluff

Stay ahead of everyone in AI.

The tightly edited AI news email engineers, founders, and investors actually open. One email. Every weekday. Five minutes to finish.

Loved by 10,000+ AI professionals
Free forever. Unsubscribe with one click.

The briefing read inside teams at

Keep reading

More from Security

ArXiv will ban authors for a year if LLMs do the work unchecked
Security

ArXiv will ban authors for a year if LLMs do the work unchecked

The preprint repository's computer science chair Thomas Dietterich set a one-strike rule for papers showing hallucinated citations or stray LLM commentary.

Jaeden Schafer4 min read
arXiv will issue one-year bans for AI-generated slop in preprints
Security

arXiv will issue one-year bans for AI-generated slop in preprints

Authors caught submitting AI-generated errors, fake citations, or misleading content face a 12-month ban and a permanent peer-review requirement.

Jaeden Schafer4 min read
Amazon is cutting up rare books to feed its AI models
Security

Amazon is cutting up rare books to feed its AI models

A tracker planted by 404 Media traced a rare book to Amazon's VGT3 facility in Las Vegas, where spines are cut and pages scanned.

Jaeden Schafer4 min read