Imparting Interpretability to Word Embeddings while Preserving Semantic Structure

Senel, Lutfi Kerem; Utlu, Ihsan; Şahinuç, Furkan; Ozaktas, Haldun M.; Koç, Aykut

doi:10.1017/S1351324920000315

Computer Science > Computation and Language

arXiv:1807.07279 (cs)

[Submitted on 19 Jul 2018 (v1), last revised 2 Jul 2020 (this version, v4)]

Title:Imparting Interpretability to Word Embeddings while Preserving Semantic Structure

Authors:Lutfi Kerem Senel, Ihsan Utlu, Furkan Şahinuç, Haldun M. Ozaktas, Aykut Koç

View PDF

Abstract:As an ubiquitous method in natural language processing, word embeddings are extensively employed to map semantic properties of words into a dense vector representation. They capture semantic and syntactic relations among words but the vectors corresponding to the words are only meaningful relative to each other. Neither the vector nor its dimensions have any absolute, interpretable meaning. We introduce an additive modification to the objective function of the embedding learning algorithm that encourages the embedding vectors of words that are semantically related to a predefined concept to take larger values along a specified dimension, while leaving the original semantic learning mechanism mostly unaffected. In other words, we align words that are already determined to be related, along predefined concepts. Therefore, we impart interpretability to the word embedding by assigning meaning to its vector dimensions. The predefined concepts are derived from an external lexical resource, which in this paper is chosen as Roget's Thesaurus. We observe that alignment along the chosen concepts is not limited to words in the Thesaurus and extends to other related words as well. We quantify the extent of interpretability and assignment of meaning from our experimental results. Manual human evaluation results have also been presented to further verify that the proposed method increases interpretability. We also demonstrate the preservation of semantic coherence of the resulting vector space by using word-analogy and word-similarity tests. These tests show that the interpretability-imparted word embeddings that are obtained by the proposed framework do not sacrifice performances in common benchmark tests.

Comments:	14 pages, 5 figures
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1807.07279 [cs.CL]
	(or arXiv:1807.07279v4 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1807.07279
Journal reference:	Natural Language Engineering, 1-26, 2020
Related DOI:	https://doi.org/10.1017/S1351324920000315

Submission history

From: Aykut Koc [view email]
[v1] Thu, 19 Jul 2018 08:14:59 UTC (293 KB)
[v2] Wed, 13 Feb 2019 13:27:55 UTC (289 KB)
[v3] Fri, 14 Feb 2020 11:41:00 UTC (339 KB)
[v4] Thu, 2 Jul 2020 11:31:59 UTC (338 KB)

Computer Science > Computation and Language

Title:Imparting Interpretability to Word Embeddings while Preserving Semantic Structure

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Imparting Interpretability to Word Embeddings while Preserving Semantic Structure

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators