Can unreviewed preprints shape scientific debate before peer review?
Explores whether papers posted to preprint servers influence research discussions and policy decisions despite lacking peer review, and what happens when institutions later challenge their reliability.
On 2025-05-16 the MIT Department of Economics published two statements about a preprint, "Artificial Intelligence, Scientific Discovery, and Product Innovation," posted to arXiv in November 2024. The first is a letter from MIT's Committee on Discipline to arXiv. After a confidential internal review, it says MIT "has no confidence in the provenance, reliability or validity of the data and has no confidence in the veracity of the research contained in the paper," and asks that the paper be marked as withdrawn. The second is a public statement. It says the paper "is already known and discussed extensively in the literature on AI and science, even though it has not been published in any refereed journal," and that MIT is speaking because, "even in its non-published form, the paper is having an impact on discussions and projections about the effects of AI on science." Its conclusion is that "the findings reported in this paper should not be relied on in academic or public discussions of these topics."
The mechanism is procedural. The excerpt notes that preprints "by definition, have not yet undergone peer review," and that only an arXiv author can submit a withdrawal request. MIT says it directed the author to submit one, that the author had not done so, and so it asked arXiv directly, and it also wrote to The Quarterly Journal of Economics, where the paper had been submitted. The letter says the review was based on allegations about "certain aspects of this paper" and that including the paper on arXiv "may violate arXiv's Code of Conduct." MIT's stated reason for going public is a formal step "it could take to mitigate the effects of misconduct," given the paper's prominence. Student privacy law and MIT policy keep the review's outcome confidential, so what outsiders receive is the institution's judgment about reliability, not the findings.
The nearest note on AI-generated research argues that journals and arXiv "can neither scale nor quality-control" that output, and proposes a venue built around closed-loop review. This case is that gap seen after the fact: arXiv hosted a paper it did not check, the paper became a reference point before any review, and the correction ran through MIT's letters, with the author's withdrawal still pending when they were written. The ICML randomized study, Does banning LLM use in peer review change review outcomes?, measures what peer review does with LLM rules; this excerpt concerns the stretch before review, where an unreviewed preprint shaped projections. The narrative-paper note's case for artifacts that keep the evidence behind each claim addresses the same provenance question MIT raises, though this excerpt says nothing about artifacts.
What the excerpt does not establish is the substance. It gives no detail of the review, the data, or the alleged problems, and it offers no evidence on the paper's findings either way. This note therefore does not treat the paper as right or wrong, and it does not cite the paper's findings as evidence for anything. What the excerpt does establish is narrower: an unreviewed preprint can shape a field's projections before anyone checks it, and when an institution later says the work should not be relied on, outsiders can only weigh that judgment. The implication is that unreviewed AI-and-science claims deserve provisional treatment until their status is settled, and that a statement like MIT's is a claim about reliability to be read as one, not a verdict.
Inquiring lines that read this note 23
This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.
How do hallucinated citations emerge in AI scholarly output? Can AI systems perform peer review as effectively as humans?- Why does publish-or-perish incentivize quantity over quality in research?
- What makes disruptive scientific work harder to publish and recognize?
- What effects do preprint servers have on scientific consensus formation?
- Can institutional statements alone correct misconceptions from unreviewed papers?
- Can automated systems scale peer review faster than human moderators?
- Should citation counts serve as the primary measure of research impact?
- Do citation counts better capture scientific quality than publication venue tiers?
- Do preprint servers have tools to detect hidden text in submitted manuscripts?
- Why do authors submit manuscripts to venues beyond their reach?
- What role do conference organizers play in accepting problematic articles?
- Does the form of a paper still matter if the process behind it changes?
- Do shortened peer review timelines correlate with lower quality publications?
- Do academic reward structures actively prevent innovation in research communication forms?
- How does document form shape what kinds of evidence social science can present?
Related concepts in this collection 3
This note in its neighbourhood — explore the map, then jump to a related concept in the list below.
Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph
-
Can automated review loops handle AI-generated research at scale?
As AI agents produce papers faster than humans can evaluate them, can a closed-loop automated review system with retrieval-augmented feedback actually improve quality and catch problems traditional peer review misses?
this case shows the arXiv quality gap that note names, after a preprint MIT now distrusts had already spread.
-
Can research papers preserve the experiments that failed?
Traditional papers compress iterative research into linear narratives, discarding failed attempts and implementation details. Could structuring papers as machine-readable packages with exploration graphs make this hidden knowledge visible and reproducible?
both target the gap between a published claim and checkable evidence, though this excerpt has no view on artifact design.
-
Does banning LLM use in peer review change review outcomes?
Can policies restricting or allowing AI tools shift how reviewers score papers and make decisions? This matters because review quality and fairness depend on consistent standards.
contrast: that study measures reviewers inside peer review; this case concerns an unreviewed preprint before any review.
Related papers in this collection 8
Papers most semantically related to this note, ranked by cosine similarity in the embedding space.
- Assuring an accurate research record
- Hidden Prompts in Manuscripts Exploit AI-Assisted Peer Review
- Screening, sorting, and the feedback cycles that imperil peer review
- Stop Automating Peer Review Without Rigorous Evaluation
- LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
- AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot
- The Emerging AI Paper-Review Arms Race: Adversarial Co-Evolution in Scholarly Publishing
- How to Find Fantastic AI Papers: Self-Rankings as a Powerful Predictor of Scientific Impact Beyond Peer Review
Original note title
MIT says it has no confidence in an AI and science preprint and asked arXiv to withdraw it — and that the unreviewed paper was already shaping debate