tlda: Tools for Language Data Analysis

Support functions and datasets to facilitate the analysis of linguistic data. The current focus is on the calculation of corpus-linguistic dispersion measures as described in Gries (2021) <doi:10.1007/978-3-030-46216-1_5> and Soenning (2025) <doi:10.3366/cor.2025.0326>. The most commonly used parts-based indices are implemented, including different formulas and modifications that are found in the literature, with the additional option to obtain frequency-adjusted scores. Dispersion scores can be computed based on individual count variables or a term-document matrix.

Package details

AuthorLukas Soenning [aut, cre, cph] (<https://orcid.org/0000-0002-2705-395X>), German Research Foundation (DFG) [fnd] (018mejw64, Grant number 548274092)
MaintainerLukas Soenning <lukas.soenning@uni-bamberg.de>
LicenseMIT + file LICENSE
Version0.1.0
URL https://github.com/lsoenning/tlda
Package repositoryView on CRAN
Installation Install the latest version of this package by entering the following in R:
install.packages("tlda")

Try the tlda package in your browser

Any scripts or data that you put into this service are public.

tlda documentation built on June 8, 2025, 11:41 a.m.