Dataset
This dataset gives the average stylometric profile of AI-generated text per language, measured from 320 samples across 5 models, 4 locales and 8 prompt types, and it is the baseline a human writer's profile is compared against.
Saying AI writing sounds generic is an assertion until there is a number for generic. These are those numbers: the centroid of AI output per language on six stylometric dimensions, measured rather than estimated, so a writer's own scores can be positioned against it.
The headline result is that the baseline is not the same in every language. Conciseness in French sits far below English, and the Japanese expressiveness figure hits the ceiling of the formula because Japanese business text uses question forms and polite markers the measure counts as expressive. A tool comparing a French writer against an English baseline gets the answer wrong by construction.
| Dimension | English | French | Spanish | Japanese |
|---|---|---|---|---|
| Sentence complexity | 65 | 75 | 71 | 62 |
| Vocabulary richness | 48 | 49 | 44 | 37 |
| Expressiveness | 76 | 74 | 59 | 100 |
| Formality | 58 | 42 | 46 | 59 |
| Consistency | 53 | 52 | 55 | 53 |
| Conciseness | 42 | 32 | 36 | 45 |
Cite this page
MyWritingTwin. “AI writing baselines by language (320 samples, 5 models, 4 locales).” Last updated 2026-09-15. https://www.mywritingtwin.com/reference/ai-baseline-divergence