AI for Science

arXiv Will Ban Researchers Who Submit Unchecked AI Slop

arXiv announced one-year bans for authors whose papers show unchecked LLM output — hallucinated references, leftover chat prompts — with later posts requiring peer review first.

arXiv Will Ban Researchers Who Submit Unchecked AI Slop — article cover
On this page7 SECTIONS
  1. The New Rule: A One-Year Ban and an “Incontrovertible Evidence” Bar
  2. What Evidence Triggers a Ban
  3. Process, Appeals, and Author Responsibility
  4. Background: October 2025 Already Tightened Review Articles
  5. Why a Preprint Platform Has No Choice
  6. What It Means in Practice for Researchers
  7. Sources

Late on Thursday, May 14, 2026, Thomas Dietterich — chair of the computer science section of arXiv — announced a new policy on X: authors whose submissions show “incontrovertible evidence that the authors did not check the results of LLM generation” will be banned from the platform for a year. After the ban lifts, any new submission must first be accepted by a reputable peer-reviewed venue. The Verge and 404 Media both covered the announcement on May 15.

This is not a social network purging spam. It is the first explicit punishment mechanism from a core piece of scientific infrastructure, a platform famous for decades of “we index, we don’t judge.” Even arXiv now concedes that the flood of AI-generated content has forced a red line.

The New Rule: A One-Year Ban and an “Incontrovertible Evidence” Bar

The trigger is deliberately narrow. The policy does not target AI-assisted writing as such — it targets papers where the evidence is beyond dispute that the authors never checked what the model produced. Dietterich’s reasoning is blunt: if the authors didn’t read the output, “this means we can’t trust anything in the paper.”

The penalty has two layers. First, a one-year ban from arXiv. Second, once the year is up, the author’s path back is narrower than everyone else’s: every future submission must clear peer review at a recognized venue before it can appear as a preprint. One violation permanently costs you the convenience of direct posting.

What Evidence Triggers a Ban

Two categories of “incontrovertible” evidence were named explicitly. The first is hallucinated references — citations to papers that do not exist. The second is LLM conversational residue left inside the manuscript: a stray “here is a 200 word summary; would you like me to make any changes?”, or a note like “the data in this table is illustrative, fill it in with the real numbers from your experiments.”

The second category is especially damning because the model itself documents that its output was never processed. These aren’t subtle failures requiring forensic detection — they are flaws visible on first read, which is precisely what makes them actionable under a moderation process.

Process, Appeals, and Author Responsibility

Procedurally, a moderator documents the problem first, and the section chair confirms it before any penalty is imposed. Dietterich told 404 Media that authors can appeal, and he stressed that the policy applies only to incontrovertible cases — a guard against punishing researchers who use AI legitimately as a writing aid.

Responsibility is stated without nuance: by signing a paper, “each author takes full responsibility for all its contents,” however it was produced. AI-generated plagiarism, errors, or misleading claims land on the authors’ record.

Background: October 2025 Already Tightened Review Articles

The ban did not come from nowhere. On October 31, 2025, arXiv’s blog announced restrictions on review articles and position papers in the computer science category, requiring them to be peer-reviewed and accepted at a conference or journal first. The stated reason: large language models had made this type of content “relatively easy to churn out on demand,” and most incoming review articles were “little more than annotated bibliographies.”

Within six months, the escalation went from restricting a genre of papers to punishing individual authors — driven by the same underlying pressure: machine-generated material pouring into the system dressed as rigorous science.

Why a Preprint Platform Has No Choice

arXiv’s value rests on being fast, open, and unreviewed — that is how it displaced paper journals on timing in the first place. But the design assumes submission cost filters for seriousness. When the cost of producing something that looks like a paper approaches zero, that assumption collapses. The peer-review system is drowning in the same flood; volunteer reviewers face submission volumes the system was never built to absorb.

A one-year ban is a heavy penalty, but compared with paywalls or full editorial review, it is the strongest tool arXiv can deploy without breaking its own openness.

What It Means in Practice for Researchers

Three concrete effects. First, AI-assisted writing is not banned — checking is now mandatory. Every reference and every data table gets verified by a human before submission. Second, teams need a last-pass review step: one co-author’s leftover LLM residue implicates every name on the paper. Third, for fields that rely on arXiv for priority claims, a one-year ban means a year absent from the scientific conversation — a cost that dwarfs the time saved by skipping verification.

Sources

AI-assisted summary compiled from the sources above, reviewed by a human before publishing.

SHAREXEMAIL