Russian-Language Electronic Fanfiction Database: Creation Principles and Quantitative Metadata Analysis
摘要
The paper is devoted to the research of fanfiction texts written and published in Russian. The study consists of creating and analyzing an electronic database which contains more than 135,000 fanfiction texts and metadata from the largest Russian-language fanfiction archive “Kniga Fanfikov” (ficbook.net). The data were retrieved through the web-scraping method and processed using Python. The database can subsequently be used for linguistic, literary, cultural, and sociological analysis, as well as implemented as a training sample for machine learning models.