Paper Search
Found 3 papers
-
Scientific production in the era of Large Language Models
Keigo Kusumegi, Xinyu Yang, Paul Ginsparg, Mathijs de Vaan, Toby Stuart, Yian Yin
January 19, 2026
cs.DL arXiv: 2601.13187Large Language Models (LLMs) are rapidly reshaping scientific research. We analyze these changes in multiple, large-scale
large language models scientific production llm adoption productivity increase writing complexity citation diversity ai detection algorithm event study models differences-in-differences analysis logistic regression poisson regression flesch reading ease score gpt-3.5 arxiv biorxiv ssrn iclr peer review openalex semantic scholar bing chat gpt-4 non-native english speakers quality signals literature discoveryphysics.soc-ph cs.AI cs.DL cs.CY -
LLM hallucinations in the wild: Large-scale evidence from non-existent citations
Zhenyue Zhao, Yihe Wang, Toby Stuart, Mathijs De Vaan, Paul Ginsparg, Yian Yin
May 08, 2026
cs.DL arXiv: 2605.07723Large language models (LLMs) are known to generate plausible but false information across a wide range of contexts, yet the real-world magnitude and consequences of this
large language models hallucination citations scientific literature arxiv biorxiv ssrn pubmed central non-existent references audit detection verification pipeline elasticsearch semantic scholar openalex google scholar llm-assisted writing authorship credit allocation gender bias peer review moderation knowledge production ai safety misinformationphysics.soc-ph cs.AI cs.DL cs.CY -
A robust association between LLM use and scientific productivity: Assessing stopping-time selection
Xinyu Yang, Keigo Kusumegi, Paul Ginsparg, Mathijs de Vaan, Toby Stuart, Yian Yin
July 31, 2026
cs.DL arXiv: 2607.28968Renault, Bergeaud, and Bosquet (hereafter RBB) argue that dating LLM adoption as the first month in which an author's abstract is flagged
llm large language models scientific productivity stopping-time selection event study difference-in-differences causal inference placebo test adoption timing detector flag rate productivity measurement artificial intelligence research output bias correction robustness check pre-chatgpt placebo intensity specification rank-based measurement conservative control group before-and-after comparison replication rebuttal methodology econometrics science of sciencecs.AI cs.DL cs.CY