Embedding Imputation with Grounded Language Information

Ziyi Yang, Chenguang Zhu, Vin Sachidananda, Eric Darve


Abstract
Due to the ubiquitous use of embeddings as input representations for a wide range of natural language tasks, imputation of embeddings for rare and unseen words is a critical problem in language processing. Embedding imputation involves learning representations for rare or unseen words during the training of an embedding model, often in a post-hoc manner. In this paper, we propose an approach for embedding imputation which uses grounded information in the form of a knowledge graph. This is in contrast to existing approaches which typically make use of vector space properties or subword information. We propose an online method to construct a graph from grounded information and design an algorithm to map from the resulting graphical structure to the space of the pre-trained embeddings. Finally, we evaluate our approach on a range of rare and unseen word tasks across various domains and show that our model can learn better representations. For example, on the Card-660 task our method improves Pearson’s and Spearman’s correlation coefficients upon the state-of-the-art by 11% and 17.8% respectively using GloVe embeddings.
Anthology ID:
P19-1326
Volume:
Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics
Month:
July
Year:
2019
Address:
Florence, Italy
Editors:
Anna Korhonen, David Traum, Lluís Màrquez
Venue:
ACL
SIG:
Publisher:
Association for Computational Linguistics
Note:
Pages:
3356–3361
Language:
URL:
https://aclanthology.org/P19-1326/
DOI:
10.18653/v1/P19-1326
Bibkey:
Cite (ACL):
Ziyi Yang, Chenguang Zhu, Vin Sachidananda, and Eric Darve. 2019. Embedding Imputation with Grounded Language Information. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, pages 3356–3361, Florence, Italy. Association for Computational Linguistics.
Cite (Informal):
Embedding Imputation with Grounded Language Information (Yang et al., ACL 2019)
Copy Citation:
PDF:
https://aclanthology.org/P19-1326.pdf
Code
 ziyi-yang/KG2Vec
Data
CARD-660ConceptNet