π₯ Explore this trending post from Hacker News π
π Category:
π‘ Main takeaway:

An AI-detection tool developed by Pangram labs found that peer reviewers are increasingly using chatbots to draft responses to authors.Credit: breakermaximus/iStock via Getty
What can researchers do if they suspect that their manuscripts have been peer reviewed using artificial intelligence (AI)? Dozens of academics have raised concerns on social media about manuscripts and peer reviews submitted to the organizers of next yearβs International Conference on Learning Representations (ICLR), an annual gathering of specialists in machine learning. Among other things, they flagged hallucinated citations and suspiciously long and vague feedback on their work.
Graham Neubig, an AI researcher at Carnegie Mellon University in Pittsburgh, Pennsylvania, was one of those who received peer reviews that seemed to have been produced using large language models (LLMs). The reports, he says, were βvery verbose with lots of bullet pointsβ and requested analyses that were not βthe standard statistical analyses that reviewers ask for in typical AI or machine-learning papers.β
But Neubig needed help proving that the reports were AI-generated. So, he posted on X (formerly Twitter) and offered a reward for anyone who could scan all the conference submissions and their peer reviews for AI-generated text. The next day, he got a response from Max Spero, chief executive of Pangram Labs in New York City, which develops tools to detect AI-generated text. Pangram screened all 19,490 studies and 75,800 peer reviews submitted for ICLR 2026, which will take place in Rio de Janeiro, Brazil, in April. Neubig and more than 11,000 other AI researchers will be attending.
Pangramβs analysis revealed that around 21% of the ICLR peer reviews were fully AI-generated, and more than half contained signs of AI use. The findings were posted online by Pangram Labs. βPeople were suspicious, but they didnβt have any concrete proof,β says Spero. βOver the course of 12 hours, we wrote some code to parse out all of the text content from these paper submissions,β he adds.
The conference organizers say they will now use automated tools to assess whether submissions and peer reviews breached policies on using AI in submissions and peer reviews. This is the first time that the conference has faced this issue at scale, says Bharath Hariharan, a computer scientist at Cornell University in Ithaca, New York, and senior programme chair for ICLR 2026. βAfter we go through all this process β¦ that will give us a better notion of trust.β
AI-written peer review
The Pangram team used one of its own tools, which predicts whether text is generated or edited by LLMs. Pangramβs analysis flagged 15,899 peer reviews that were fully AI-generated. But it also identified many manuscripts that had been submitted to the conference with suspected cases of AI-generated text: 199 manuscripts (1%) were found to be fully AI-generated; 61% of submissions were mostly human-written; but 9% contained more than 50% AI-generated text.
Pangram described the model in a preprint1, which it submitted to ICLR 2026. Of the four peer reviews received for the manuscript, one was flagged as fully AI-generated and another as lightly AI-edited, the teamβs analysis found.

AI is transforming peer review β and many scientists are worried
For many researchers who received peer reviews for their submissions to ICLR, the Pangram analysis confirmed what they had suspected. Desmond Elliott, a computer scientist at the University of Copenhagen, says that one of three reviews he received seemed to have missed βthe point of the paperβ. His PhD student who led the work suspected that the review was generated by LLMs, because it mentioned numerical results from the manuscript that were incorrect and contained odd expressions.
When Pangram released its findings, Elliott adds, βthe first thing I did was I typed in the title of our paper because I wanted to know whether my studentβs gut instinct was correctβ. The suspect peer review, which Pangramβs analysis flagged as fully AI-generated, gave the manuscript the lowest rating, leaving it βon the borderline between accept and rejectβ, says Elliott. βIt’s deeply frustratingβ.
Repercussions
π¬ What do you think?
#οΈβ£ #Major #conference #flooded #peer #reviews #written #fully
π Posted on 1764435620
