01
By genre
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
| genre | n | peak position | late − early | peaks > 0.5 sd | final level |
|---|---|---|---|---|---|
| adventure | 11 | 0.85 | 0.51 | 2.00 | 0.48 |
| comic | 11 | 0.59 | 0.33 | 3.00 | 0.12 |
| detective | 10 | 0.79 | 0.92 | 3.00 | 0.30 |
| ghost | 13 | 0.92 | 1.19 | 3.00 | 0.85 |
| horror | 18 | 0.95 | 1.11 | 2.00 | 1.00 |
| literary | 15 | 0.87 | 0.81 | 2.00 | 0.42 |
| speculative | 14 | 0.82 | 0.78 | 3.00 | 0.56 |
02
By era and by length
Publication era is confounded with genre and author in a corpus this size; treat differences as descriptive. Length bands test whether the 0→1 normalization hides a real dependence on absolute length.
Era
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
Permutation p = 0.696; between-era share of variance 1.8%.
Story length
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
Permutation p = 0.237; between-band share 2.6%.
03
By author
Authors with at least three stories in the corpus. Author identity is the strongest nuisance variable for a text-based measure: register, sentence rhythm and vocabulary are all author-level.
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
04
Clustering curve shapes — against a null
K-means on z-scored curves for k = 2…6. A silhouette is only meaningful next to its null: the same clustering on within-story shuffled curves. Where the observed silhouette does not clear the null’s 95th percentile, the “archetypes” are what k-means finds in noise.
| k | silhouette | null mean | null 95th pct | beats null? |
|---|---|---|---|---|
| 2 | 0.106 | 0.034 | 0.040 | yes |
| 3 | 0.108 | 0.031 | 0.038 | yes |
| 4 | 0.101 | 0.027 | 0.031 | yes |
| 5 | 0.096 | 0.024 | 0.032 | yes |
| 6 | 0.084 | 0.019 | 0.027 | yes |
k = 3 centroids (shown regardless of whether they beat the null)
Genre mix: comic 5 · literary 4 · speculative 4 · horror 3 · adventure 3 · detective 2
05
Same question, other measures
The genre picture should not depend on one operationalization. Below, the by-genre curves for each alternative measure with its permutation p-value.
Composite lexical proxy · permutation p = 0.817
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
GPT-2 surprisal · permutation p = 0.011
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
Threat lexicon · permutation p = 0.842
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
Uncertainty · permutation p = 0.098
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
LLM intensity · permutation p = 0.371
Select one group to see its 95% bootstrap band over stories. Curves are z-scored within story before averaging, so height differences between stories are removed and only shape remains.
06
Extremes
The earliest and latest peaks in the corpus on the primary measure.