arXiv 因 AI 灌水论文激增限制投稿数量
ArXiv Is Rate Limiting Submissions Because It Can't Keep Up with AI Slop
arXiv 宣布对投稿实施速率限制,每位研究者每个自然月最多提交两篇,同时最多有三篇处于活跃状态的投稿。arXiv 称 AI 让灌水变得过于容易,志愿者审核团队不堪重负,投稿量从 2016 年 9 月的 9,869 篇增至 2024 年 9 月的 20,569 篇,2026 年 9 月达到 40,363 篇,两年翻倍,计算机科学类目增长六倍。
arXiv, an open-access repository where researchers publish preprint academic research, has announced it will rate limit submissions because it has been inundated with AI-written papers. A researcher can now only send in two pieces per calendar month and have a total of three active submissions at any given time. In a blog post about the change, arXiv said AI has made it too easy to spam the publisher’s inbox and its volunteer moderation team is overwhelmed.
“The heart of arXiv’s process for detecting and rejecting low-quality papers is our many volunteer moderators,” Thomas Dietterich, an Oregon State University professor emeritus and the chair of arXiv’s editorial advisory council said in a statement. “We are so grateful that they donate their time and expertise every day. However, a relatively small proportion of authors are submitting a large number of low-quality papers and consuming a disproportionate fraction of the moderators’ time. This is unfair to authors who continue to submit quality papers — their papers can be delayed for days or weeks as a result.”
According to arXiv, it received 9,869 submissions in September 2016, 20,569 submissions in September 2024, and 40,363 submissions in September 2026. Submissions on the site have doubled in the last two years and have sextupled in the computer science category.
AI is driving this surge in multiple ways. Researchers are using LLM to parse massive data sets, write code, and design portions of experiments; some are also using AI to write the papers. “arXiv policy permits authors to employ AI as a tool in support of their research as long as it is disclosed and the research meets arXiv’s standards of scholarly interest and advances the field of research,” arXiv said in its rate limiting blog. “However, many of the submissions that arXiv is now receiving do not satisfy these requirements.”
arXiv said it’s seen an increase in “thin papers of narrow scope” as well as salami papers — where researchers write multiple articles based on one study. “There is also a marked increase in dense, AI-written papers. AI tools are making it easy for authors to flood arXiv and other repositories with these low-value papers,” it said.
Rate limiting was already part of arXiv’s policy, but it was originally left to each individual moderator’s discretion. “This policy update strives to fairly distribute the incredibly valuable work of our volunteer moderators who check that submissions meet these standards,” the publisher said.
This new policy is another part of a long battle the publication is having with AI-powered slop and spam. A little less than a year ago, the archive stopped accepting review articles and position papers in the computer science category and cited low effort LLM-written articles as the reason. In January it started requiring new authors to receive an endorsement from an existing author in the system. Previously the publisher accepted an academic email address as the sole qualifier. In May, arXiv took things a step farther and said it would ban researchers who submitted AI slop from the site for a year.
About the author
Matthew Gault is a writer covering weird tech, nuclear war, and video games. He’s worked for Reuters, Motherboard, and the New York Times.
来源:Hacker News · AI · 404media.co