Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. NeurIPS 2026 desk-rejects papers with hallucinated…

NeurIPS 2026 desk-rejects papers with hallucinated references and runs a randomized LLM-assisted reviewing experiment

★★★after cutoffpolicy-safetyNeurIPSGoogleconfidence: medium

NeurIPS 2026 treats hallucinated citations as a Code of Conduct violation. An area chair reported on Aug 18, 2026 that 5 of the 8 submissions in his batch had two or more fabricated references and would likely be desk-rejected. The same cycle brought AI into the process officially: an opt-in Google Gemini "Paper Assistant Tool" gave authors feedback before submission, and an IRB-approved experiment randomly assigned reviewers to no LLM help, open-ended help or structured help from an LLM assistant built into OpenReview.

Key facts

What happened

NeurIPS 2026, the largest ML conference, dealt with AI on both sides of peer review. Before submission, authors could opt in to one automated critique per paper from Google's Gemini-based Paper Assistant Tool. During review, NeurIPS ran a randomized, IRB-approved experiment. Volunteer reviewers on opted-in papers got no LLM help, open-ended help or structured help from an assistant built into OpenReview, and blinded area chairs rated the resulting reviews. Any other reviewer LLM use stayed banned.

NeurIPS also enforced its rule against fabricated citations. On Aug 18 area chair Danish Pruthi reported that most papers in his batch had several hallucinated references and faced desk rejection, while one fabricated citation alone was not enough.

Why it matters

Fabricated references are an easy-to-check sign of unchecked LLM writing, and the top ML venues have started enforcing against them. At the same time they are testing, under controlled conditions, whether LLM help improves reviews. The ICML 2026 experiment suggests that policy alone barely changes outcomes and that many reviewers ignore restrictions.

Changelog

  • 2026-10-02: created (leads: strictcite.com / Zvi AI #188 trail)

Related posts (2)

Related events

  1. NeurIPS 2026: 28% of position-track submissions score 100% AI-written, and 178 are desk-rejected ★★
  2. ICLR 2026 review crisis: 21% of peer reviews flagged fully AI-written, and an OpenReview bug exposes reviewer identities ★★★
  3. arXiv will ban authors for a year if they post unchecked LLM-generated content ★★★

Sources (8)

id: 2026-08-18-neurips-2026-hallucinated-references-ai-reviewing · updated 2026-10-02 · open in the interactive timeline