| spokenBNC2014_metadata | R Documentation |
This dataset provides some metadata for the speakers in the Spoken BNC2014 (Love et al. 2017), including information on age, gender, and the total number of word tokens contributed to the corpus.
spokenBNC2014_metadata
spokenBNC2014_metadataA data frame with 668 rows and 6 columns:
Speaker ID (e.g. "S0001", "S0002")
Age group, based on the BNC1994 scheme ("0-14", "15-24", "25-34", "35-44", "45-59", "60+", "Unknown")
Speaker gender ("Female" vs. "Male")
Age of speaker; if actual age is not available, imputed based on age_group and age_bin
Number of word tokens the speaker contributed to the corpus
Age group, based on the BNC2014 scheme ("0-9", "10-19", "20-29", "30-39", "40-49", "50-59", "60-69", "70+")
Love, Robbie, Claire Dembry, Andrew Hardie, Vaclav Brezina & Tony McEnery. 2017. The Spoken BNC2014: Designing and building a spoken corpus of everyday conversations. International Journal of Corpus Linguistics, 22(3), 319–344.
Add the following code to your website.
For more information on customizing the embed code, read Embedding Snippets.