Can peer review gates stop the flood of AI-generated surveys?
arXiv CS now requires prior journal or conference peer review for survey and position papers. The question is whether this upstream gate actually reduces low-quality submissions and restores the moderation workload to manageable levels.
arXiv's computer science category has changed its moderation practice for review (survey) articles and position papers. Such a paper must now be accepted by a journal or conference after complete peer review, and the author must submit documentation of that review. Papers without it "will be likely to be rejected and not appear on arXiv." The stated reason is "the unmanageable influx of review articles and position papers to arXiv CS." The post also corrects the record: these types were never officially accepted content types, and had been admitted "at moderator discretion" because the few received were of high quality. The volume figure is arXiv's own: "we now receive hundreds of review articles every month." It is not an external measurement.
The mechanism is the one the post gives. Large language models made this content "relatively easy to churn out on demand," and arXiv says "the majority of the review articles we receive are little more than annotated bibliographies, with no substantial discussion of open research issues." That majority claim is the moderators' judgment, offered without a sample. Volunteer moderators lack "the time or bandwidth" to vet hundreds of such papers, so arXiv hands quality control to refereed venues, which "conduct in-depth review to assure quality, evidential support of opinions, and scholarly value." The post excludes workshop review from that standard. Research papers, including social-science work in cs.CY, are not covered by the change.
Set against the nearest notes, this is a different answer to the same volume pressure than the one Can human review keep pace with AI-accelerated research generation? argues for. That note says review must itself become automated once generation is AI-accelerated. arXiv does the reverse: it adds a human-refereed gate upstream and keeps its own moderation human. It does confirm the limit that Can automated review loops handle AI-generated research at scale? names, since arXiv says it "does not have the resources to conduct this quality-control in-house" for content types it does not accept. The gate also leans on the venues' peer review. The ICML experiment in Does banning LLM use in peer review change review outcomes? found that a substantial share of reviewers broke the LLM-use rules they were given, so the arXiv rule inherits that weakness.
The excerpt does not establish that the rule works. It gives no before-and-after submission counts, no rejection rate, and no evidence that the refereed venues are free of the flood it describes. "Likely to be rejected" is a hedged enforcement statement. The rule also targets content type, not AI authorship: the excerpt never says that using AI disqualifies a paper. So the evidence supports a narrower claim than the post's framing. arXiv changed its rule and gave its reason, but nothing here shows the change restored quality. Anyone citing it as a fix for LLM-written surveys should carry that gap.
Inquiring lines that read this note 1
This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.
What human oversight must AI research systems have?Related concepts in this collection 4
This note in its neighbourhood — explore the map, then jump to a related concept in the list below.
Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph
-
Can automated review loops handle AI-generated research at scale?
As AI agents produce papers faster than humans can evaluate them, can a closed-loop automated review system with retrieval-augmented feedback actually improve quality and catch problems traditional peer review misses?
arXiv's own admission that it cannot run quality control in-house is a concrete instance of the limit this note names.
-
Can human review keep pace with AI-accelerated research generation?
As AI systems generate hypotheses, code, and proofs faster than humans can verify them, does the bottleneck at peer review force verification itself to become automated? What governance structures enable this transition safely?
contrast: that note automates review under generation pressure; arXiv adds a human-refereed gate upstream instead.
-
Does banning LLM use in peer review change review outcomes?
Can policies restricting or allowing AI tools shift how reviewers score papers and make decisions? This matters because review quality and fairness depend on consistent standards.
qualifies the gate: arXiv relies on venue peer review, whose LLM-use rules reviewers broke in that experiment.
-
Where is most LLM-generated content actually appearing in computer science?
Review papers show higher LLM-generated shares than other papers, but the raw volume tells a different story. This note explores which paper types actually contain most generated content by count.
Qualifies the flood claim: LLM share is higher in review papers, but generated non-review papers outnumber generated reviews almost sixfold
Related papers in this collection 8
Papers most semantically related to this note, ranked by cosine similarity in the embedding space.
- Attention Authors: Updated Practice for Review Articles and Position Papers in arXiv CS Category
- LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
- Assuring an accurate research record
- Use and Effects of LLMs in Peer Review: A Randomized Experiment and Survey at ICML 2026
- Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
- Position: The AI Conference Peer Review Crisis Demands Author Feedback and Reviewer Rewards
- Screening, sorting, and the feedback cycles that imperil peer review
- Stop Automating Peer Review Without Rigorous Evaluation
Original note title
arXiv CS requires prior peer review for review and position papers because their influx became unmanageable, which LLMs made easy to churn out