Support functions and datasets to facilitate the analysis of linguistic data. The current focus is on the calculation of corpus-linguistic dispersion measures as described in Gries (2021) <doi:10.1007/978-3-030-46216-1_5> and Soenning (2025) <doi:10.3366/cor.2025.0326>. The most commonly used parts-based indices are implemented, including different formulas and modifications that are found in the literature, with the additional option to obtain frequency-adjusted scores. Dispersion scores can be computed based on individual count variables or a term-document matrix.
Package details |
|
|---|---|
| Author | Lukas Soenning [aut, cre, cph] (<https://orcid.org/0000-0002-2705-395X>), German Research Foundation (DFG) [fnd] (018mejw64, Grant number 548274092) |
| Maintainer | Lukas Soenning <lukas.soenning@uni-bamberg.de> |
| License | MIT + file LICENSE |
| Version | 0.1.0 |
| URL | https://github.com/lsoenning/tlda |
| Package repository | View on CRAN |
| Installation |
Install the latest version of this package by entering the following in R:
|
Any scripts or data that you put into this service are public.
Add the following code to your website.
For more information on customizing the embed code, read Embedding Snippets.