| spokenBNC1994_metadata | R Documentation |
This dataset provides some metadata for speakers in the demographically sampled part of the Spoken BNC1994 (Crowdy 1995), including information on age, gender, and the total number of word tokens contributed to the corpus.
spokenBNC1994_metadata
spokenBNC1994_metadataA data frame with 1,017 rows and 7 columns:
Speaker ID (e.g. "PS002", "PS003")
Age group, based on the BNC1994 scheme ("0-14", "15-24", "25-34", "35-44", "45-59", "60+", "Unknown")
Speaker gender ("Female" vs. "Male")
Age of speaker; if actual age is not available, imputed based on age_group and age_bin
Number of word tokens the speaker contributed to the corpus
Age group, based on the BNC2014 scheme ("0-9", "10-19", "20-29", "30-39", "40-49", "50-59", "60-69", "70+")
Crowdy, Steve. 1995. The BNC spoken corpus. In Geoffrey Leech, Greg Myers & Jenny Thomas (eds.), Spoken English on Computer: Transcription, Mark-Up and Annotation, 224–234. Harlow: Longman.
Add the following code to your website.
For more information on customizing the embed code, read Embedding Snippets.