Word sense disambiguation by semantic inference

This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words f...

Full description

Saved in:
Bibliographic Details
Published in2017 International Conference on Behavioral, Economic, Socio-cultural Computing (BESC) pp. 1 - 6
Main Authors Xinda Wang, Xuri Tang, Weiguang Qu, Min Gu
Format Conference Proceeding
LanguageEnglish
Published IEEE 01.10.2017
Subjects
Online AccessGet full text

Cover

Loading…
More Information
Summary:This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words for the target word in a sentence, builds substitute constructs by replacing the target word with substitute words, uses large-scale dependency parsed corpora to calculate the likelihood of the substitute constructs, and then obtain the best substitute word which help specify the sense of the target word in the sentence. Experiments with WordNet 2.1 and the corpora English Gigawords on the lexical sample task in SemEval-2007 show that the algorithm achieves the-state-of-art accuracy for both nouns and verbs, which are 3-5 percent higher than the best unsupervised system in SemEval-2007, given the condition that the knowledge source provides sufficient information.
DOI:10.1109/BESC.2017.8256391