Comparison of Deep-Neural-Network-Based Models for Estimating Distributed Representations of Compound Words
An Dao, Natthawut Kertkeidkachorn, Ryutaro Ichise
- 发表年份
- 2021
- 引用次数
- 4
摘要
Word embeddings or word vectors have become fundamental in language processing techniques, especially deep learning approaches. Although many languages have compound words (e.g., “robot arm” and “maple leaf”), such words have not received much attention from researchers. Most research on compound word embeddings considered only two-word compounds; there has been little detailed analysis on the learning representations of arbitrary-length compound words. This paper discusses the necessity for learning-based approaches for estimating the distributed representations of compound words instead of a simple average of the representations of constituents. An evaluation of two downstream tasks confirms the effectiveness of compositional models in encoding useful information into vector spaces. The experimental results suggest that complex architectures such as long short-term memory, gated recurrent units, and transformers learn better representations for long entities, whereas simpler models such as recurrent neural networks are more applicable for downstream tasks where there are only short compounds (two or three words in length), as in the noun compound interpretation task.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991
A new optimizer using particle swarm theory
R.C. Eberhart, James Kennedy
2002