SYNTHESIS NOTE
Topics›Domain Specialization›this note

Can unreviewed preprints shape scientific debate before peer review?

Explores whether papers posted to preprint servers influence research discussions and policy decisions despite lacking peer review, and what happens when institutions later challenge their reliability.

Synthesis note · 2026-10-06 · sourced from Domain Specialization

On 2025-05-16 the MIT Department of Economics published two statements about a preprint, "Artificial Intelligence, Scientific Discovery, and Product Innovation," posted to arXiv in November 2024. The first is a letter from MIT's Committee on Discipline to arXiv. After a confidential internal review, it says MIT "has no confidence in the provenance, reliability or validity of the data and has no confidence in the veracity of the research contained in the paper," and asks that the paper be marked as withdrawn. The second is a public statement. It says the paper "is already known and discussed extensively in the literature on AI and science, even though it has not been published in any refereed journal," and that MIT is speaking because, "even in its non-published form, the paper is having an impact on discussions and projections about the effects of AI on science." Its conclusion is that "the findings reported in this paper should not be relied on in academic or public discussions of these topics."

The mechanism is procedural. The excerpt notes that preprints "by definition, have not yet undergone peer review," and that only an arXiv author can submit a withdrawal request. MIT says it directed the author to submit one, that the author had not done so, and so it asked arXiv directly, and it also wrote to The Quarterly Journal of Economics, where the paper had been submitted. The letter says the review was based on allegations about "certain aspects of this paper" and that including the paper on arXiv "may violate arXiv's Code of Conduct." MIT's stated reason for going public is a formal step "it could take to mitigate the effects of misconduct," given the paper's prominence. Student privacy law and MIT policy keep the review's outcome confidential, so what outsiders receive is the institution's judgment about reliability, not the findings.

The nearest note on AI-generated research argues that journals and arXiv "can neither scale nor quality-control" that output, and proposes a venue built around closed-loop review. This case is that gap seen after the fact: arXiv hosted a paper it did not check, the paper became a reference point before any review, and the correction ran through MIT's letters, with the author's withdrawal still pending when they were written. The ICML randomized study, Does banning LLM use in peer review change review outcomes?, measures what peer review does with LLM rules; this excerpt concerns the stretch before review, where an unreviewed preprint shaped projections. The narrative-paper note's case for artifacts that keep the evidence behind each claim addresses the same provenance question MIT raises, though this excerpt says nothing about artifacts.

What the excerpt does not establish is the substance. It gives no detail of the review, the data, or the alleged problems, and it offers no evidence on the paper's findings either way. This note therefore does not treat the paper as right or wrong, and it does not cite the paper's findings as evidence for anything. What the excerpt does establish is narrower: an unreviewed preprint can shape a field's projections before anyone checks it, and when an institution later says the work should not be relied on, outsiders can only weigh that judgment. The implication is that unreviewed AI-and-science claims deserve provisional treatment until their status is settled, and that a statement like MIT's is a claim about reliability to be read as one, not a verdict.

Inquiring lines that read this note 23

This note is a source for these research framings, grouped by the broader line of inquiry each explores. Scan the bold lines of inquiry; follow any specific question forward.

How do hallucinated citations emerge in AI scholarly output? Can AI systems perform peer review as effectively as humans? Does AI-assisted research sacrifice exploration breadth for productivity gains? Can artificial systems establish authority in domains requiring expert judgment? Do restrictions on reviewer LLM use actually shape peer review behavior? What gaps exist between benchmark performance and real deployment outcomes? What human oversight must AI research systems have? What governance mechanisms can effectively constrain widely deployed AI systems? Can we trust AI-generated mathematical proofs without understanding them? Why do LLM research ideation systems generate novelty but lack diversity?

Related concepts in this collection 3

This note in its neighbourhood — explore the map, then jump to a related concept in the list below.

Concept map
15 direct connections · 63 in 2-hop network ·medium cluster Open in graph ↗

Click a node to walk · click center to open · click Open in graph to see this note in the full knowledge graph

your link semantically near linked from elsewhere

Related papers in this collection 8

Papers most semantically related to this note, ranked by cosine similarity in the embedding space.

Original note title

MIT says it has no confidence in an AI and science preprint and asked arXiv to withdraw it — and that the unreviewed paper was already shaping debate