Towards a Zero-Data, Controllable, Adaptive Dialog System

Väth, Dirk; Vanderlyn, Lindsey; Vu, Ngoc Thang

Computer Science > Computation and Language

arXiv:2403.17582 (cs)

[Submitted on 26 Mar 2024]

Title:Towards a Zero-Data, Controllable, Adaptive Dialog System

Authors:Dirk Väth, Lindsey Vanderlyn, Ngoc Thang Vu

View PDF HTML (experimental)

Abstract:Conversational Tree Search (Väth et al., 2023) is a recent approach to controllable dialog systems, where domain experts shape the behavior of a Reinforcement Learning agent through a dialog tree. The agent learns to efficiently navigate this tree, while adapting to information needs, e.g., domain familiarity, of different users. However, the need for additional training data hinders deployment in new domains. To address this, we explore approaches to generate this data directly from dialog trees. We improve the original approach, and show that agents trained on synthetic data can achieve comparable dialog success to models trained on human data, both when using a commercial Large Language Model for generation, or when using a smaller open-source model, running on a single GPU. We further demonstrate the scalability of our approach by collecting and testing on two new datasets: ONBOARD, a new domain helping foreign residents moving to a new city, and the medical domain DIAGNOSE, a subset of Wikipedia articles related to scalp and head symptoms. Finally, we perform human testing, where no statistically significant differences were found in either objective or subjective measures between models trained on human and generated data.

Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2403.17582 [cs.CL]
	(or arXiv:2403.17582v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2403.17582

Submission history

From: Dirk Väth [view email]
[v1] Tue, 26 Mar 2024 10:45:11 UTC (3,376 KB)

Computer Science > Computation and Language

Title:Towards a Zero-Data, Controllable, Adaptive Dialog System

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Towards a Zero-Data, Controllable, Adaptive Dialog System

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators