Automatic Selection of Context Configurations for Improved Class-Specific Word Representations

Vulić, Ivan; Schwartz, Roy; Rappoport, Ari; Reichart, Roi; Korhonen, Anna

Computer Science > Computation and Language

arXiv:1608.05528 (cs)

[Submitted on 19 Aug 2016 (v1), last revised 12 Jun 2017 (this version, v3)]

Title:Automatic Selection of Context Configurations for Improved Class-Specific Word Representations

Authors:Ivan Vulić, Roy Schwartz, Ari Rappoport, Roi Reichart, Anna Korhonen

View PDF

Abstract:This paper is concerned with identifying contexts useful for training word representation models for different word classes such as adjectives (A), verbs (V), and nouns (N). We introduce a simple yet effective framework for an automatic selection of class-specific context configurations. We construct a context configuration space based on universal dependency relations between words, and efficiently search this space with an adapted beam search algorithm. In word similarity tasks for each word class, we show that our framework is both effective and efficient. Particularly, it improves the Spearman's rho correlation with human scores on SimLex-999 over the best previously proposed class-specific contexts by 6 (A), 6 (V) and 5 (N) rho points. With our selected context configurations, we train on only 14% (A), 26.2% (V), and 33.6% (N) of all dependency-based contexts, resulting in a reduced training time. Our results generalise: we show that the configurations our algorithm learns for one English training setup outperform previously proposed context types in another training setup for English. Moreover, basing the configuration space on universal dependencies, it is possible to transfer the learned configurations to German and Italian. We also demonstrate improved per-class results over other context types in these two languages.

Comments:	CoNLL 2017
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:1608.05528 [cs.CL]
	(or arXiv:1608.05528v3 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.1608.05528

Submission history

From: Ivan Vulić [view email]
[v1] Fri, 19 Aug 2016 08:30:35 UTC (159 KB)
[v2] Wed, 7 Jun 2017 09:26:29 UTC (160 KB)
[v3] Mon, 12 Jun 2017 13:11:53 UTC (153 KB)

Computer Science > Computation and Language

Title:Automatic Selection of Context Configurations for Improved Class-Specific Word Representations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Automatic Selection of Context Configurations for Improved Class-Specific Word Representations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators