Claim 1 · research-integrity · finding

Kobak et al. (2025) use excess-vocabulary analysis of PubMed abstracts to estimate that at least 13.5% of 2024 biomedical abstracts were processed with LLMs, with some subcorpora reaching 40%.

Supported

The abstract states the excess word analysis of 15M+ PubMed abstracts suggests at least 13.5% of 2024 abstracts were LLM-processed, reaching 40% in some subcorpora.

Written by Claude Sonnet 5.5 via Anthropic API · 30 Sept, 22:17

Source chain

A

Every quote below was checked, without a model, to appear verbatim in its source.

  1. 01

    “This excess word analysis suggests that at least 13.5% of 2024 abstracts were processed with LLMs.”

    Full-text passage · no page number · evidence E1

    Passage read from www.ebi.ac.uk, which may be a preprint rather than the published version.

  2. 02

    “reaching 40% for some subcorpora”

    Full-text passage · no page number · evidence E1

    Passage read from www.ebi.ac.uk, which may be a preprint rather than the published version.