On Possibilities and Methods of Analysis of Thematic Expressions in Spoken Texts
21 déc. 2019
À propos de cet article
Publié en ligne: 21 déc. 2019
Pages: 469 - 480
DOI: https://doi.org/10.2478/jazcas-2019-0075
Mots clés
© 2019 Petr Pořízka, published by Sciendo
This work is licensed under the Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 License.
The treatise focuses on mutual comparison of three methods of detection of prominent text units (prominent in relation to the contents of the text). The methods are: 1) analysis of key words based on comparison of source and referential corpora, 2) thematic concentration and h-point, and 3) the TF*IDF method. We try to thematize their pros and cons and, using the results of the carried out analyses, propose the optimal method for the extraction of thematic words from the spoken texts the frequency structure of which differs distinctly from the frequency structure of written texts.