How we measured AI writing across arXiv, and where the measurement breaks
- AI
- Science
- Education
- Trust & Safety
The post analyzes 12,750 arXiv papers and applies a custom detector meant to estimate how much academic writing now looks machine-generated. The author says the detector was tuned so pre-ChatGPT papers only trip it about 0.4% of the time, then reports a steep rise through 2026, with computer science far above math. The post is careful to frame this as a statistical estimate, not proof that any one paper was written by a model, and it explicitly includes AI-assisted editing in what might get flagged.
Treat these results as a population-level style signal, not evidence about any specific paper or author. If AI-assisted writing is becoming normal in research, the harder problem for teams, reviewers, and publishers is filtering for truth, novelty, and reproducibility rather than policing prose alone.
-
unslop.run
- Discuss on HN