Learning English with Peppa Pig

Nikolaus, Mitja; Alishahi, Afra; Chrupała, Grzegorz

doi:10.1162/tacl_a_00498

Computer Science > Computation and Language

arXiv:2202.12917 (cs)

[Submitted on 25 Feb 2022 (v1), last revised 27 May 2022 (this version, v2)]

Title:Learning English with Peppa Pig

Authors:Mitja Nikolaus, Afra Alishahi, Grzegorz Chrupała

View PDF

Abstract:Recent computational models of the acquisition of spoken language via grounding in perception exploit associations between the spoken and visual modalities and learn to represent speech and visual data in a joint vector space. A major unresolved issue from the point of ecological validity is the training data, typically consisting of images or videos paired with spoken descriptions of what is depicted. Such a setup guarantees an unrealistically strong correlation between speech and the visual data. In the real world the coupling between the linguistic and the visual modality is loose, and often confounded by correlations with non-semantic aspects of the speech signal. Here we address this shortcoming by using a dataset based on the children's cartoon Peppa Pig. We train a simple bi-modal architecture on the portion of the data consisting of dialog between characters, and evaluate on segments containing descriptive narrations. Despite the weak and confounded signal in this training data our model succeeds at learning aspects of the visual semantics of spoken language.

Comments:	Accepted to TACL
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Audio and Speech Processing (eess.AS); Image and Video Processing (eess.IV)
Cite as:	arXiv:2202.12917 [cs.CL]
	(or arXiv:2202.12917v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2202.12917
Related DOI:	https://doi.org/10.1162/tacl_a_00498

Submission history

From: Grzegorz Chrupała [view email]
[v1] Fri, 25 Feb 2022 19:14:35 UTC (528 KB)
[v2] Fri, 27 May 2022 17:54:11 UTC (494 KB)

Computer Science > Computation and Language

Title:Learning English with Peppa Pig

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Learning English with Peppa Pig

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators