media creative tools
The Model Trace Reached More Than Half of Astronomy Papers
A Bayesian study estimates 54% of 2025 astro-ph papers carried language-model markers, while 0.81% disclosed their use.
Summary
A Bayesian study estimates 54% of 2025 astro-ph papers carried language-model markers, while 0.81% disclosed their use.
The analysis counts distinctive vocabulary across 207,111 astro-ph papers from 2015 to mid-2026, calibrating unassisted prose with pre-2020 papers and assisted prose with 392 disclosures. Its central 2025 estimate is 54%, with an eight-point statistical interval and a one-sided systematic range as wide as 26 points because today’s no-model background cannot be observed directly. The estimate remains at least 36% under tested choices. Only 0.81% of 2025 papers disclosed model use—roughly one declaration per 66 papers carrying a trace. Vocabulary markers are indirect and adaptive, so the result estimates assistance rather than proving it paper by paper.
Why it matters
A Bayesian study estimates 54% of 2025 astro-ph papers carried language-model markers, while 0.81% disclosed their use.
Limits and context
- Its central 2025 estimate is 54%, with an eight-point statistical interval and a one-sided systematic range as wide as 26 points because today’s no-model background cannot be observed directly.
- Only 0.81% of 2025 papers disclosed model use—roughly one declaration per 66 papers carrying a trace.
Key claims
A Bayesian study estimates 54% of 2025 astro-ph papers carried language-model markers, while 0.81% disclosed their use.
Qualification: Its central 2025 estimate is 54%, with an eight-point statistical interval and a one-sided systematic range as wide as 26 points because today’s no-model background cannot be observed directly.
Evidence: source-2026-09-11-016
Sources
- arXiv preprint 2609.10664arXiv · primary research
Corrections
No corrections have been recorded for this story.