Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Lautaro Quiroz

Splitting Compounds by Semantic Analogy

Sep 15, 2015

Joachim Daiber, Lautaro Quiroz, Roger Wechsler, Stella Frank

Figure 1 for Splitting Compounds by Semantic Analogy

Figure 2 for Splitting Compounds by Semantic Analogy

Figure 3 for Splitting Compounds by Semantic Analogy

Figure 4 for Splitting Compounds by Semantic Analogy

Abstract:Compounding is a highly productive word-formation process in some languages that is often problematic for natural language processing applications. In this paper, we investigate whether distributional semantics in the form of word embeddings can enable a deeper, i.e., more knowledge-rich, processing of compounds than the standard string-based methods. We present an unsupervised approach that exploits regularities in the semantic vector space (based on analogies such as "bookshop is to shop as bookshelf is to shelf") to produce compound analyses of high quality. A subsequent compound splitting algorithm based on these analyses is highly effective, particularly for ambiguous compounds. German to English machine translation experiments show that this semantic analogy-based compound splitter leads to better translations than a commonly used frequency-based method.

* Proceedings of the 1st Deep Machine Translation Workshop. Prague, Czech Republic. 2015

Via

Access Paper or Ask Questions