ArXiv bans AI-generated papers: How science adapts to AI revolution
Keywords
Artificial IntelligenceBy Jean Piot
August 12, 2026

ArXiv (pronounced "ar-kive", like archive in English : the X stands for the Greek letter "Chi") is an open-access preprint platform for scientific documents. The documents published here undergo no peer review, a traditionally mandatory step in the scientific publication process. While the system may be less robust, it is significantly faster than the traditional process.
Warning
On October 31, 2025, ArXiv’s editorial team announced on its blog a policy change: prior to publication, submissions in the computer science domain would now be reviewed by a reading committee instead of being solely validated by a moderator. This decision was made in response to the surge in AI-generated documents lacking any added value. Currently, ArXiv receives thousands of articles per month. The platform, established in 2001, lacks the infrastructure to handle all submissions.
The ruling
On May 14, 2026, Thomas Dietterich, a computer science professor at the University of Oregon and ArXiv’s computer science section president, published this warning:
"To all ArXiv authors: our Code of Conduct states that every author signing a submission assumes full responsibility for its content—regardless of how the text was created. The author remains liable even if the submitted text is generated by an AI and contains inappropriate expressions, plagiarism, errors, incorrect references, or misleading content. We have recently clarified the penalties in this regard. If a submission contains irrefutable evidence that the author did not verify what the AI wrote, we cannot trust them. They will be banned from ArXiv for one year, and subsequently, all their submissions will require expert committee validation."
How can we tell if the author did not proofread their document?
According to Thomas Dietterich, submitted texts may include, for example:
- made-up references.
- known false or unjustified facts.
- comments such as 'this is an example, would you like me to modify something?' or 'replace the table data with your own data', clear signs of copying and pasting an AI-generated response.
As reported by The Next Web, this decision only concerns authors who made blatant careless mistakes. But it says nothing about the use of AI as an aid in writing.
The limitations of this decision
There will nevertheless remain on the platform articles written by LLMs (*Large Language Models* — AI systems capable of generating text), with a veneer of scientific credibility but little substance. There will also be articles with plausible but incorrect conclusions, as well as documents presenting completely fabricated data. Removing them will require not only time, but also peer-review committees. And perhaps AI, even though AI-based detection of text written by AI is currently not 100% reliable. Obviously, the goal is not to ban artificial intelligence altogether. But it is essential to regulate its use... which is much easier said than done.


