All-words Word Sense Disambiguation for Russian Using Automatically Generated Text Collection

Angelina B.; Loukachevitch N.

All-words Word Sense Disambiguation for Russian Using Automatically Generated Text Collection

Angelina B.; Loukachevitch N.

URI: https://dspace.kpfu.ru/xmlui/handle/net/162182

Дата: 2020

Аннотации:

© 2020 Bolshina Angelina et al., published by Sciendo 2020. The limited amount of the sense annotated data is a big challenge for the word sense disambiguation task. As a solution to this problem, we propose an algorithm of automatic generation and labelling of the training collections based on the monosemous relatives concept. In this article we explore the limits of this algorithm: we employ it to harvest training collections for all ambiguous nouns, verbs and adjectives presented in RuWordNet thesaurus and then evaluate the quality of the obtained collections. We demonstrate that our approach can create high-quality labelled collections with almost full-coverage of the RuWordNet polysemous words. Furthermore, we show that our method can be applied to the Word-in-Context task.

Показать полную информацию

Файлы в этом документе

Имя: SCOPUS13119702-20 ...

Размер: 54.48Kb

Формат: PDF

Открыть

Данный элемент включен в следующие коллекции

Публикации сотрудников КФУ Scopus [24551]
Коллекция содержит публикации сотрудников Казанского федерального (до 2010 года Казанского государственного) университета, проиндексированные в БД Scopus, начиная с 1970г.