| vowel_cohort | R Documentation |
A simulated cohort for demonstrating the ranking protocol of
rank_contrasts(): twelve speakers, two vowel categories
("ih" and "eh"), and two acoustic features (F1 and F2, in
Hz). Each speaker was built to exercise one part of the protocol:
Gaussian categories whose centroid gap grows from 25 to 200 Hz, 100 tokens per vowel: a graded ordering that both Jensen-Shannon distance and Pillai recover.
The same centroids for both vowels, but a bimodal
"eh" (two variants 190 Hz apart in F1 and 440 Hz apart in F2),
100 tokens per vowel. A mean-based measure sees no contrast while the
distributions barely overlap: the planted Pillai / \sqrt{JSD}
disagreement the agreement flag should catch.
Fully separated categories, 100 tokens per vowel:
\sqrt{JSD} at the ceiling.
60 tokens per vowel: ranking by \sqrt{JSD} is licensed
at two dimensions (floor 50) but the flag is not readable (floor 100).
40 tokens per vowel: below the two-dimensional rank floor, so the speaker is ordered by Pillai.
vowel_cohort
A data frame with 2200 rows and 4 columns:
Character; speaker identifier "spk01" to
"spk12".
Character; vowel category, "ih" or "eh".
Numeric; first formant frequency in Hz.
Numeric; second formant frequency in Hz.
Simulated with a fixed seed; the generating script is
data-raw/vowel_cohort.R in the package repository.
rank_contrasts(), inspect_contrast().
head(vowel_cohort)
table(vowel_cohort$speaker, vowel_cohort$vowel)
Add the following code to your website.
For more information on customizing the embed code, read Embedding Snippets.