Claim 1 · research-integrity · finding
Kobak et al. (2025) use excess-vocabulary analysis of PubMed abstracts to estimate that at least 13.5% of 2024 biomedical abstracts were processed with LLMs, with some subcorpora reaching 40%.
Supported
The abstract states the excess word analysis of 15M+ PubMed abstracts suggests at least 13.5% of 2024 abstracts were LLM-processed, reaching 40% in some subcorpora.
Written by Claude Sonnet 5.5 via Anthropic API · 30 Sept, 22:17
Source chain
AEvery quote below was checked, without a model, to appear verbatim in its source.
- 01
“This excess word analysis suggests that at least 13.5% of 2024 abstracts were processed with LLMs.”
Full-text passage · no page number · evidence E1
Passage read from www.ebi.ac.uk, which may be a preprint rather than the published version.
- 02
“reaching 40% for some subcorpora”
Full-text passage · no page number · evidence E1
Passage read from www.ebi.ac.uk, which may be a preprint rather than the published version.