Observed Signal · May 17, 2026 · Policy Update · Source: t3n · Impact: 2/5 · Sentiment: Neutral
arXiv tightens rules for AI-generated research papers
arXiv, the open preprint repository formerly hosted by Cornell University and now transitioning to an independent nonprofit, has introduced stricter rules governing submissions that rely on large language models (LLMs). The platform requires first authors to have an endorsement from an established author and places responsibility on authors to verify any AI-generated content. arXiv warns that submissions containing clear evidence that authors did not check LLM outputs — for example, fabricated citations — may trigger sanctions: a one-year ban and a requirement that subsequent submissions be accepted by a recognized peer‑review platform. Moderators and subject-area chairs must confirm violations before penalties are applied, and appeals are possible. The change follows broader concerns about increasing fabricated references in biomedical literature and widespread use of AI tools in universities.
A policy update from a major academic preprint repository signals growing scrutiny of LLM-generated content and could influence detection tools, research practices, and trust in AI-assisted outputs—relevant to vendors and platforms working with generative AI, but not a direct AdTech platform policy change.
Track arXiv Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- arXiv introduced stricter guidelines for submissions involving large language models (LLMs).
- First authors must now provide an endorsement from an established author for submissions.
- arXiv states that if a submission contains irrefutable evidence authors did not verify LLM-generated results (e.g., fabricated citations), sanctions can include a one-year ban and conditional acceptance rules.
- Moderators must report issues and subject-area chairs must confirm evidence before penalties are imposed; appeals are allowed.
- External reports note rising fabricated citations in biomedical literature and high student use of AI tools (a 2025 survey at Hochschule Darmstadt found 92% of students used tools like ChatGPT at least occasionally).
Connected Companies & Entities
4 Entities mappedOntology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Study Finds ~147K Fake Citations from LLM Hallucinations
A new study analyzed 111 million references across 2.5 million scientific papers and identified 146,900 fabricated or untraceable citations, attributing a steep rise in such false references to the widespread use of large language models (LLMs). The fake citations appeared across preprint repositories including arXiv, bioRxiv, SSRN and PubMed Central. Researchers from Cornell University and the University of California conducted the analysis. In response, arXiv has tightened submission rules—requiring stronger author verification and recommending sanctions (including potential one‑year bans) for works that show authors did not verify LLM-generated content. Scientists and platform leaders warned that AI hallucinations dilute the scientific record and undermine trust in research literature.
Wikipedia bans AI‑generated article text
Wikipedia updated its editorial policy on March 26, 2026, prohibiting editors from using large language models (LLMs) to generate or rewrite article content. The change tightens prior guidance that only discouraged using LLMs to create new articles from scratch. The policy change was approved by editors in a community vote (40 to 2). It still permits limited AI assistance: editors may use LLMs to suggest basic copyedits for their own writing and incorporate suggestions after human review, provided the LLM does not introduce new content. The policy emphasizes caution because LLMs can alter meaning or produce text not supported by cited sources.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
