Monday, August 3, 2026

Policy & Regulation

ArXiv to ban authors for one year over unverified AI content

ArXiv will issue a one-year ban to authors who submit research containing unverified AI-generated content, marking a strict crackdown on the use of LLMs in scientific papers.

ArXiv to ban authors for one year over unverified AI content

ArXiv, an open repository for preprint research—defined as research papers posted to a repository before undergoing peer review—is implementing a new one-strike policy. Under this rule, the repository will ban authors for a 1-year period if their submissions contain evidence that they did not verify AI-generated content. While papers on the platform are posted before they are peer-reviewed, the repository serves as a primary channel for circulating research in fields such as computer science and mathematics, and the site has become a source of data on trends in scientific research.

The policy update was posted on Thursday by Thomas Dietterich, the chair of arXiv’s computer science section. Dietterich explained that “if a submission contains incontrovertible evidence that the authors did not check the results of LLM generation, this means we can’t trust anything in the paper.” An LLM, or Large Language Model, is an AI system used to generate text. The policy is not an outright ban on these models. Instead, it demands that authors take full responsibility for their submissions, regardless of how the contents are generated. Evidence of a failure to verify, such as hallucinated references or direct comments to or from an LLM, will trigger the penalty. According to Dietterich, the penalty consists of a 1-year ban from arXiv, followed by a requirement that subsequent submissions must first be accepted by a reputable peer-reviewed venue. In an interview with the media outlet 404 Media, Dietterich noted that moderators must flag the issue and section chairs must confirm the evidence before imposing the penalty, and authors will be able to appeal.

The policy addresses a broader concern regarding AI slop and fabricated citations in scientific literature. Peer-reviewed research indicates that fabricated citations are likely increasing in biomedical research. The policy change also coincides with structural shifts at the repository. After being hosted by Cornell for more than 20 years, arXiv is transitioning to become an independent nonprofit organization, a transition that should allow it to raise more money to address issues like AI slop.

Why it matters

This policy represents a significant escalation in the scientific community’s effort to maintain research integrity against the proliferation of low-quality, AI-generated content. By penalizing unverified AI use, the repository establishes a clear standard of accountability for researchers utilizing automated tools.