THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS

Standard

THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS: book chapter. / Zenkov, Andrei Viacheslavovich ; Zenkov, Eugene Viacheslavovich; Zenkov, Miroslav Andreevich.
РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ: монография. Петрозаводск: Международный центр научного партнерства «Новая Наука», 2021. p. 153-201.

Research output: Chapter in Book/Report/Conference proceeding › Chapter › peer-review

Harvard

Zenkov, AV , Zenkov, EV & Zenkov, MA 2021, THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS: book chapter. in РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ: монография. Международный центр научного партнерства «Новая Наука», Петрозаводск, pp. 153-201.

APA

Zenkov, A. V., Zenkov, E. V., & Zenkov, M. A. (2021). THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS: book chapter. In РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ: монография (pp. 153-201). Международный центр научного партнерства «Новая Наука».

Vancouver

Zenkov AV , Zenkov EV, Zenkov MA. THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS: book chapter. In РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ: монография. Петрозаводск: Международный центр научного партнерства «Новая Наука». 2021. p. 153-201

Author

Zenkov, Andrei Viacheslavovich ; Zenkov, Eugene Viacheslavovich ; Zenkov, Miroslav Andreevich. / THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS : book chapter. РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ: монография. Петрозаводск : Международный центр научного партнерства «Новая Наука», 2021. pp. 153-201

BibTeX

@inbook{2116d3285edc4f488c8aaee4c916902a,

title = "THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS: book chapter",

abstract = "We present some results obtained in the framework of the project “The novel method of text attribution based on the numerals statistics” supported by a grant from the Russian Foundation for Basic Research, project No. 19-012-00199A. We suggest two approaches to the statistical analysis of texts, both based on the study of numerals occurrence in texts. The 1st approach is related to the study of the frequency distribution of various leading digits of numerals occurring in the text. These frequencies are unequal: the digit 1 is strongly dominating; usually, the incidence of subsequent digits is monotonically decreasing. The frequencies of occurrence of the digit 1, as well as, to a lesser extent, the digits 2 and 3, are usually a characteristic author's style feature, manifested in all (sufficiently long) literary texts of any author. This approach is convenient for testing whether a group of texts has common authorship: the latter is dubious if the frequency distributions are sufficiently different. The 2nd approach is the extension of the first one and requires the study of the frequency distribution of numerals themselves (not their leading digits). The approach yields non-trivial information about the author, stylistic and genre peculiarities of the texts and is suited for the advanced stylometric analysis. The proposed approaches are illustrated by examples of computer analysis of the literary texts in Russian, Czech, Lithuanian, English, and Turkish.",

author = "Zenkov, {Andrei Viacheslavovich} and Zenkov, {Eugene Viacheslavovich} and Zenkov, {Miroslav Andreevich}",

year = "2021",

language = "English",

isbn = "978-5-00174-293-7",

pages = "153--201",

booktitle = "РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ",

publisher = "Международный центр научного партнерства «Новая Наука»",

address = "Russian Federation",

}

RIS

TY - CHAP

T1 - THE NOVEL METHOD OF TEXT ATTRIBUTION BASED ON THE NUMERALS STATISTICS: A SURVEY OF RESULTS

T2 - book chapter

AU - Zenkov, Andrei Viacheslavovich

AU - Zenkov, Eugene Viacheslavovich

AU - Zenkov, Miroslav Andreevich

PY - 2021

Y1 - 2021

N2 - We present some results obtained in the framework of the project “The novel method of text attribution based on the numerals statistics” supported by a grant from the Russian Foundation for Basic Research, project No. 19-012-00199A. We suggest two approaches to the statistical analysis of texts, both based on the study of numerals occurrence in texts. The 1st approach is related to the study of the frequency distribution of various leading digits of numerals occurring in the text. These frequencies are unequal: the digit 1 is strongly dominating; usually, the incidence of subsequent digits is monotonically decreasing. The frequencies of occurrence of the digit 1, as well as, to a lesser extent, the digits 2 and 3, are usually a characteristic author's style feature, manifested in all (sufficiently long) literary texts of any author. This approach is convenient for testing whether a group of texts has common authorship: the latter is dubious if the frequency distributions are sufficiently different. The 2nd approach is the extension of the first one and requires the study of the frequency distribution of numerals themselves (not their leading digits). The approach yields non-trivial information about the author, stylistic and genre peculiarities of the texts and is suited for the advanced stylometric analysis. The proposed approaches are illustrated by examples of computer analysis of the literary texts in Russian, Czech, Lithuanian, English, and Turkish.

AB - We present some results obtained in the framework of the project “The novel method of text attribution based on the numerals statistics” supported by a grant from the Russian Foundation for Basic Research, project No. 19-012-00199A. We suggest two approaches to the statistical analysis of texts, both based on the study of numerals occurrence in texts. The 1st approach is related to the study of the frequency distribution of various leading digits of numerals occurring in the text. These frequencies are unequal: the digit 1 is strongly dominating; usually, the incidence of subsequent digits is monotonically decreasing. The frequencies of occurrence of the digit 1, as well as, to a lesser extent, the digits 2 and 3, are usually a characteristic author's style feature, manifested in all (sufficiently long) literary texts of any author. This approach is convenient for testing whether a group of texts has common authorship: the latter is dubious if the frequency distributions are sufficiently different. The 2nd approach is the extension of the first one and requires the study of the frequency distribution of numerals themselves (not their leading digits). The approach yields non-trivial information about the author, stylistic and genre peculiarities of the texts and is suited for the advanced stylometric analysis. The proposed approaches are illustrated by examples of computer analysis of the literary texts in Russian, Czech, Lithuanian, English, and Turkish.

UR - https://www.elibrary.ru/item.asp?id=46444847

M3 - Chapter

SN - 978-5-00174-293-7

SP - 153

EP - 201

BT - РАЗВИТИЕ ОБЩЕСТВА И НАУКИ В УСЛОВИЯХ ЦИФРОВОЙ ЭКОНОМИКИ

PB - Международный центр научного партнерства «Новая Наука»

CY - Петрозаводск

ER -

ID: 23765508