Skip to content
World clockEU--:--UK--:--USA--:--CN--:--PLDEFRIT中文EN

portal about AI and technologyevents · analysis · interviews · technical background

Search
LIVE
›

arXiv caps submissions as AI floods preprint servers with low-quality papers

On 1 October, the arXiv preprint server capped submissions to two per person per month, a direct response to a surge in AI-generated content that has doubled its monthly intake since 2024.

AI & modelsNewsRachel NwosuPublished: 4 October 20264 min readSources 10
arXiv caps submissions as AI floods preprint servers with low-quality papers

The move comes as research institutions grapple with a logistics crisis. According to The Register, arXiv received 40,363 submissions in September 2026, almost double the 20,569 received in September 2024. This volume is more than four times the 9,869 submissions recorded in September 2016.

Kat Boboris announced the policy. She noted that the latest month's submissions generated almost 9,000 support tickets for staff. She observed a marked increase in dense, AI-written papers, describing a trend of 'salami' slicing where single works are broken into smaller, lower-value pieces.

The quality signal is fading

Daniel Lemire, a professor at the University of Quebec, noted that the repository has seen submissions doubling every two years, a pace he calls unsustainable for human moderators. The system relies on volunteers to filter out garbage and spam, but the volume has overwhelmed their capacity. In related news, Google recently stopped taking new bug reports in its open-source bounty program for similar reasons of influx.

'AI tools are making it easy for authors to flood arXiv and other repositories with these low-value papers,' Boboris said, as cited by The Register.

This flood is not merely a quantity issue; it is a quality crisis. A new collaboration between Seoul National University and the University of Minnesota proposes a method to detect 'scientific slop.' Their paper, titled 'Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers,' suggests that conventional AI-text detectors miss structural weaknesses. The system looks beyond the text to assess whether claims are properly supported by arguments, evidence, and citations.

Each part of a paper can look plausible in isolation, yet the connections between sections may be broken.

Integrity and incentive problems

The rise of AI-generated content has exacerbated existing incentive problems in academic research. A blog post by msoos.org argues that the outcome of academic research has drifted from its original goal of advancing scientific understanding. The author notes that once a paper is published and promotions secured, the correctness of the results often becomes a secondary concern.

'It’s not your problem. It’s someone else’s problem,' the post states, reflecting a sentiment that researchers are not interested in learning that their papers are wrong. This culture creates an environment where incorrect or misleading results remain unchallenged, potentially harming the common understanding of science.

Junior researchers are increasingly influenced by these incentives. The msoos.org author recalls a PhD student who, when told their evaluation approach was incorrect, did not follow up. This contrasts with older generations who felt 'deep shame' when their methods were flawed and immediately fixed their tools.

The security research angle

While the academic community debates the quality of AI-generated theory, the security sector is leveraging AI to scale vulnerability research. Caleb Gross, a security researcher, presented a method for using Large Language Models to prioritize attack surfaces. His tool, Slice, successfully reproduced the discovery of a use-after-free vulnerability in the Linux kernel SMB server for approximately $3 per run.

Another tool, Raink, ranks datasets with O(N) complexity, allowing researchers to prioritize thousands of GitHub repositories or kernel subsystems. This efficiency suggests that while AI may be flooding preprint servers with low-quality theory, it is also enabling more cost-effective and scalable practical security work.

Broader implications for scientific infrastructure

The pressure on preprint servers is part of a larger shift in how research is conducted and distributed. The Turing Tree tracker shows that top AI researchers like Yoshua Bengio and Geoffrey Hinton continue to accumulate citations, with Bengio reaching 1,164,407 citations as of October 2026. However, the rapid increase in low-quality submissions threatens to dilute the value of these metrics.

arXiv had previously implemented a one-year ban for unchecked AI content in May 2026. The new two-submission cap is a further retrenchment to preserve the integrity of the repository. For now, the non-profit status of arXiv means there is little financial incentive to enforce logins or mandatory site visits, but the sustainability of the volunteer model is increasingly doubtful.

As the volume of AI-generated content continues to rise, the scientific community may need to develop new standards for validation. The current system, designed for a slower era of human-driven research, is struggling to keep up with the speed and scale of machine-generated output.

Comments 0

Sources

10
  1. 01Fighting AI slop in science papersEN
  2. 02Research paper overload: submissions capped at two a monthEN
  3. 03Incentives in Academic ResearchEN
  4. 04O(N) the Money: Scaling Vulnerability Research with LLMs (2025)EN
  5. 05Top 50 AI researchers by citationsEN
  6. 06When Does Automating AI Research Produce Explosive Growth?EN
  7. 07What's the future for pure math research in the age of AI?EN
  8. 08PL research is dead, the age of PL exploration is just beginningEN
  9. 09wizard-engine: Research WebAssembly EngineEN
  10. 10Tracking vulnerabilities that credit the Anthropic research teamEN

All figures and quotations in this text come from the sources listed below.

Content prepared by the editorial team with AI assistance.

Rachel Nwosu

Rachel Nwosu

AI, models and technology

Rachel Nwosu covers AI, models and technology for FLASH24, working from public model documentation, benchmark releases and repository histories rather than press summaries, and she skips announcements that arrive without reproducible numbers. She checks training-data claims against dataset cards and reruns reported metrics where code is available. She spends much of her week interviewing researchers and engineers, tracking model launch calendars, and comparing vendor benchmarks with independent evaluations. Outside the desk she runs 3D printers, restores old computers, and tests how models learn from internet junk. She does not publish benchmark figures she cannot trace to a source.

Newsroom →

Comments

0
  1. No comments yet — be the first.

Write a comment

Comments are public. We do not publish abuse, spam or advertising.