Word sense disambiguation by semantic inference

This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words f...

Full description

Saved in:

Bibliographic Details
Published in	2017 International Conference on Behavioral, Economic, Socio-cultural Computing (BESC) pp. 1 - 6
Main Authors	Xinda Wang, Xuri Tang, Weiguang Qu, Min Gu
Format	Conference Proceeding
Language	English
Published	IEEE 01.10.2017
Subjects	Inference algorithms Information retrieval semantic inference Semantics Thesauri Training Training data unsupervised word sense disambiguation
Online Access	Get full text

Cover

Loading…

More Information
Summary:	This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words for the target word in a sentence, builds substitute constructs by replacing the target word with substitute words, uses large-scale dependency parsed corpora to calculate the likelihood of the substitute constructs, and then obtain the best substitute word which help specify the sense of the target word in the sentence. Experiments with WordNet 2.1 and the corpora English Gigawords on the lexical sample task in SemEval-2007 show that the algorithm achieves the-state-of-art accuracy for both nouns and verbs, which are 3-5 percent higher than the best unsupervised system in SemEval-2007, given the condition that the knowledge source provides sufficient information.
DOI:	10.1109/BESC.2017.8256391