A New Set of Linguistic Resources for Ukrainian
摘要
We have constructed a Ukrainian set of linguistic resources that have allowed us to construct various NLP applications on Ukrainian, including information retrieval and extraction, morphological, syntactic semantic and statistical analysis, spell checking, and machine translation. Our goal was to develop a reliable tool that would allow students and teachers of the Ukrainian language to explore simple texts, as well to allow researchers in the social sciences to analyze their own corpora of Ukrainian texts. We will first review the various existing NLP software applications that can process Ukrainian texts, their functionalities, and their performance. We then describe the linguistic resources we have developed, and finally compare the results produced by both approaches.