Learning Disentangled Intent Representations for Zero-shot Intent Detection

Si, Qingyi; Liu, Yuanxin; Fu, Peng; Li, Jiangnan; Lin, Zheng; Wang, Weiping

Computer Science > Computation and Language

arXiv:2012.01721v1 (cs)

[Submitted on 3 Dec 2020 (this version), latest version 8 Jun 2021 (v2)]

Title:Learning Disentangled Intent Representations for Zero-shot Intent Detection

Authors:Qingyi Si, Yuanxin Liu, Peng Fu, Jiangnan Li, Zheng Lin, Weiping Wang

View PDF

Abstract:Zero-shot intent detection (ZSID) aims to deal with the continuously emerging intents without annotated training data. However, existing ZSID systems suffer from two limitations: 1) They are not good at modeling the relationship between seen and unseen intents, when the label names are given in the form of raw phrases or sentences. 2) They cannot effectively recognize unseen intents under the generalized intent detection (GZSID) setting. A critical factor behind these limitations is the representations of unseen intents, which cannot be learned in the training stage. To address this problem, we propose a class-transductive framework that utilizes unseen class labels to learn Disentangled Intent Representations (DIR). Specifically, we allow the model to predict unseen intents in the training stage, with the corresponding label names serving as input utterances. Under this framework, we introduce a multi-task learning objective, which encourages the model to learn the distinctions among intents, and a similarity scorer, which estimates the connections among intents more accurately based on the learned intent representations. Since the purpose of DIR is to provide better intent representations, it can be easily integrated with existing ZSID and GZSID methods. Experiments on two real-world datasets show that the proposed framework brings consistent improvement to the baseline systems, regardless of the model architectures or zero-shot learning strategies.

Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2012.01721 [cs.CL]
	(or arXiv:2012.01721v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2012.01721

Submission history

From: Qingyi Si [view email]
[v1] Thu, 3 Dec 2020 06:41:09 UTC (1,359 KB)
[v2] Tue, 8 Jun 2021 18:18:33 UTC (835 KB)

Computer Science > Computation and Language

Title:Learning Disentangled Intent Representations for Zero-shot Intent Detection

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Learning Disentangled Intent Representations for Zero-shot Intent Detection

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators