Evaluate Similarity of Requirements with Multilingual Natural Language Processing
Abstract Finding redundant requirements or semantically similar ones in previous projects is a very time-consuming task in engineering design, especially with multilingual data. Due to modern NLP it is possible to automate such tasks. In this paper we compared different multilingual embeddings model...
Saved in:
Published in | Proceedings of the Design Society Vol. 2; pp. 1511 - 1520 |
---|---|
Main Authors | , , , , |
Format | Journal Article Conference Proceeding |
Language | English |
Published |
Cambridge
Cambridge University Press
01.05.2022
|
Subjects | |
Online Access | Get full text |
Cover
Loading…
Summary: | Abstract
Finding redundant requirements or semantically similar ones in previous projects is a very time-consuming task in engineering design, especially with multilingual data. Due to modern NLP it is possible to automate such tasks. In this paper we compared different multilingual embeddings models to see which of them is the most suitable to find similar requirements in English and German. The comparison was done for both in-domain data (requirements pairs) and out-of-domain data (general sentence pairs). The most suitable model were sentence embeddings learnt with knowledge distillation. |
---|---|
ISSN: | 2732-527X 2732-527X |
DOI: | 10.1017/pds.2022.153 |