Imparting interpretability to word embeddings while preserving semantic structure

Şenel, L. K.; Utlu, İhsan; Şahinuç, Furkan; Özaktaş, Haldun M.; Koç, Aykut

Imparting interpretability to word embeddings while preserving semantic structure

buir.contributor.author	Utlu, İhsan
buir.contributor.author	Şahinuç, Furkan
buir.contributor.author	Özaktaş, Haldun M.
buir.contributor.author	Koç, Aykut
dc.citation.epage	26	en_US
dc.citation.spage	1	en_US
dc.contributor.author	Şenel, L. K.
dc.contributor.author	Utlu, İhsan
dc.contributor.author	Şahinuç, Furkan
dc.contributor.author	Özaktaş, Haldun M.
dc.contributor.author	Koç, Aykut
dc.date.accessioned	2021-03-08T08:11:14Z
dc.date.available	2021-03-08T08:11:14Z
dc.date.issued	2020
dc.department	Department of Electrical and Electronics Engineering	en_US
dc.department	National Magnetic Resonance Research Center (UMRAM)	en_US
dc.description.abstract	As a ubiquitous method in natural language processing, word embeddings are extensively employed to map semantic properties of words into a dense vector representation. They capture semantic and syntactic relations among words, but the vectors corresponding to the words are only meaningful relative to each other. Neither the vector nor its dimensions have any absolute, interpretable meaning. We introduce an additive modification to the objective function of the embedding learning algorithm that encourages the embedding vectors of words that are semantically related to a predefined concept to take larger values along a specified dimension, while leaving the original semantic learning mechanism mostly unaffected. In other words, we align words that are already determined to be related, along predefined concepts. Therefore, we impart interpretability to the word embedding by assigning meaning to its vector dimensions. The predefined concepts are derived from an external lexical resource, which in this paper is chosen as Roget’s Thesaurus. We observe that alignment along the chosen concepts is not limited to words in the thesaurus and extends to other related words as well. We quantify the extent of interpretability and assignment of meaning from our experimental results. Manual human evaluation results have also been presented to further verify that the proposed method increases interpretability. We also demonstrate the preservation of semantic coherence of the resulting vector space using word-analogy/word-similarity tests and a downstream task. These tests show that the interpretability-imparted word embeddings that are obtained by the proposed framework do not sacrifice performances in common benchmark tests.	en_US
dc.identifier.doi	10.1017/S1351324920000315	en_US
dc.identifier.eissn	1469-8110
dc.identifier.issn	1351-3249
dc.identifier.uri	http://hdl.handle.net/11693/75864
dc.language.iso	English	en_US
dc.publisher	Cambridge University Press	en_US
dc.relation.isversionof	https://dx.doi.org/10.1017/S1351324920000315	en_US
dc.source.title	Natural Language Engineering	en_US
dc.subject	Word embeddings	en_US
dc.subject	Interpretability	en_US
dc.subject	Computational semantics	en_US
dc.title	Imparting interpretability to word embeddings while preserving semantic structure	en_US
dc.type	Article	en_US

Files

Original bundle

Now showing 1 - 1 of 1

Name:: Imparting_interpretability_to_word_embeddings_while_preserving_semantic_structure.pdf
Size:: 1 MB
Format:: Adobe Portable Document Format
Description:: View / Download

Download

License bundle

Now showing 1 - 1 of 1

Name:: license.txt
Size:: 1.71 KB
Format:: Item-specific license agreed upon to submission
Description:

Download

Collections

Scholarly Publications - Electrical and Electronics Engineering
Scholarly Publications - UMRAM