Word sense disambiguation by semantic inference
This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words f...
Saved in:
Published in | 2017 International Conference on Behavioral, Economic, Socio-cultural Computing (BESC) pp. 1 - 6 |
---|---|
Main Authors | , , , |
Format | Conference Proceeding |
Language | English |
Published |
IEEE
01.10.2017
|
Subjects | |
Online Access | Get full text |
Cover
Loading…
Summary: | This paper proposes an algorithm for unsupervised Word Sense Disambiguation to bypass the knowledge bottleneck faced by supervised approaches. By simulating the semantic inference process performed by human language users, the algorithm makes use of a thesaurus to obtain potential substitute words for the target word in a sentence, builds substitute constructs by replacing the target word with substitute words, uses large-scale dependency parsed corpora to calculate the likelihood of the substitute constructs, and then obtain the best substitute word which help specify the sense of the target word in the sentence. Experiments with WordNet 2.1 and the corpora English Gigawords on the lexical sample task in SemEval-2007 show that the algorithm achieves the-state-of-art accuracy for both nouns and verbs, which are 3-5 percent higher than the best unsupervised system in SemEval-2007, given the condition that the knowledge source provides sufficient information. |
---|---|
DOI: | 10.1109/BESC.2017.8256391 |