Word frequency–rank relationship in tagged texts
Andrés Chacoma and
Damián H. Zanette
Physica A: Statistical Mechanics and its Applications, 2021, vol. 574, issue C
Abstract:
We analyze the frequency–rank relationship in sub-vocabularies corresponding to three different grammatical classes (nouns, verbs, and others) in a collection of literary works in English, whose words have been automatically tagged according to their grammatical role. Comparing with a null hypothesis which assumes that words belonging to each class are uniformly distributed across the frequency–ranked vocabulary of the whole work, we disclose statistically significant differences between the three classes. This results point to the fact that frequency–rank relationships may reflect linguistic features associated with grammatical function.
Keywords: Frequency–rank statistics; Grammatical function; Linguistic regularities; Language processing; Quantitative linguistics (search for similar items in EconPapers)
Date: 2021
References: View references in EconPapers View complete reference list from CitEc
Citations: View citations in EconPapers (1)
Downloads: (external link)
http://www.sciencedirect.com/science/article/pii/S0378437121002922
Full text for ScienceDirect subscribers only. Journal offers the option of making the article available online on Science direct for a fee of $3,000
Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.
Export reference: BibTeX
RIS (EndNote, ProCite, RefMan)
HTML/Text
Persistent link: https://EconPapers.repec.org/RePEc:eee:phsmap:v:574:y:2021:i:c:s0378437121002922
DOI: 10.1016/j.physa.2021.126020
Access Statistics for this article
Physica A: Statistical Mechanics and its Applications is currently edited by K. A. Dawson, J. O. Indekeu, H.E. Stanley and C. Tsallis
More articles in Physica A: Statistical Mechanics and its Applications from Elsevier
Bibliographic data for series maintained by Catherine Liu ().