AI Writers Have a Consistent Stylometric Footprint, but AI Editors Do Not
cs.CL
Submitted: 2026-08-28
Updated: 2026-09-22
Comments: EMNLP Main 2026
License: http://creativecommons.org/licenses/by/4.0/
The gist: Text generated by large language models (LLMs) has been shown to be stylometrically distinct from human-written text.
Terminology
Abstract
Text generated by large language models (LLMs) has been shown to be stylometrically distinct from human-written text. But LLMs are increasingly used not only to generate text but also to edit human writing, and it is unclear whether the two leave the same trace. We show that AI generation leaves a consistent ``stylometric footprint'': a small subset of features, primarily entropy and lexical diversity, consistently separates AI-generated text from human writing across 8 LLMs and 5 domains, while the remaining features depend heavily on the domain and generator. AI editing, however, does not reproduce the same footprint. Relative to their human-written sources, AI-edited texts show only a small increase in lexical diversity and a decrease in entropy, rather than the joint increase that characterizes AI generation. Lexical density, which contributes little to generation, instead becomes the dominant editing-associated signal. Stylometric features therefore separate AI-edited text from AI-generated text but are substantially less effective at separating it from human-written text. Our results suggest that ``AI text'' is not a single phenomenon: generation and editing leave qualitatively different stylometric traces and should be studied separately.
Sources
- GPTZero: Robust Detection of LLM-Generated Texts
- Hierarchical Neural Story Generation
- gpt-oss-120b & gpt-oss-20b Model Card
- The Llama 3 Herd of Models
- How Close is ChatGPT to Human Experts? Comparison Corpus, Evaluation, and Detection
- Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Text
- Improve LLM-based Automatic Essay Scoring with Linguistic Features
- Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature
- Gemma 3 Technical Report
- DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability Curvature
- Implementing BERT and fine-tuned RobertA to detect AI generated news by ChatGPT
- Qwen2.5 Technical Report
- A Survey of Large Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering