AI news ·
arXiv caps submissions to two per month after AI-driven research surge
arXiv now limits authors to two submissions per month after new papers nearly doubled to 40,363 in September 2026. A detection tool called SciSlopHarness correctly flags AI-generated papers 85.9% of the time by testing structural reasoning, not surface patterns.

arXiv began limiting authors to two submissions per month on October 1, a direct response to a surge in AI-generated papers that has overwhelmed the preprint server's moderation systems. September 2026 saw 40,363 new submissions-nearly double the 20,569 in September 2024 and quadruple the 9,869 a decade earlier.
The policy applies across all subject categories but targets a flood of low-quality machine-generated content that has strained peer review and editorial workflows. arXiv's moderation team confirmed the rate limit took effect immediately, with no grandfathering for frequent submitters.
What the submission numbers show
The growth trajectory is steep. Submissions rose from roughly 10,000 per month in 2016 to 20,000 by 2024, then doubled again in two years. The 2026 figures represent a compound annual growth rate far above historical norms for the repository, which has operated since 1991.
arXiv has not publicly attributed the entire increase to AI-generated content. However, the platform's statement pointed to "unsustainable submission volumes" that correlate with the widespread availability of large language models. The two-paper cap is the first blanket restriction arXiv has imposed on submission frequency.
A benchmark for detecting scientific slop
Researchers from Seoul National University and the University of Minnesota released SciSlopBench, a benchmark designed to catch what they call "scientific slop"-papers that fail on structure, argumentation, and evidence rather than just surface-level text patterns. Their detection tool, SciSlopHarness, correctly identified AI-generated papers 85.9% of the time in tests, outperforming standard AI-text detectors.
The team built the benchmark by pairing 390 AI-generated papers with human-written counterparts. They found that machine-generated papers often contain individually plausible sections that fall apart when examined for cross-sectional reasoning, citation accuracy, and logical coherence. A live demo of the tool scores papers on these dimensions, giving the researchers' own paper a 40 out of 100 on the slop scale.
"AI-generated papers often lack coherent reasoning across sections, citations, and evidence, even if individual parts appear plausible," the authors wrote. The benchmark tests for failures that traditional detectors miss because those tools focus on sentence-level statistical signatures rather than structural integrity.
Implications for researchers and institutions
The two-paper limit will affect prolific research groups that use arXiv as a primary distribution channel. Labs accustomed to posting multiple preprints per month-whether human-written or AI-assisted-now face a hard constraint. The policy does not distinguish between sole authors and large collaborations, meaning multi-institution projects must coordinate submission slots carefully.
For readers and reviewers, the cap may reduce the noise in daily arXiv feeds. But it also raises questions about whether rate-limiting addresses the root problem or simply slows the firehose. The SciSlopBench researchers argue that detection tools must become part of the submission pipeline, not just post-hoc filters.
Why this matters for writers, educators, and researchers
The arXiv cap signals that AI-generated content has reached a volume where platforms must impose blunt restrictions to maintain basic quality control. For writers and educators who teach or rely on research, this means the tools for distinguishing human scholarship from machine output are becoming as important as the research itself. SciSlopHarness and similar benchmarks offer a concrete method for evaluating papers-not by hunting for AI fingerprints, but by checking whether the argument actually holds together across sections. That skill, evaluating structural coherence, is one that working professionals in writing and education can apply immediately, with or without detection software.