Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA

Ciosici, Manuel R.; Cecil, Joe; Hedges, Alex; Lee, Dong-Ho; Freedman, Marjorie; Weischedel, Ralph

Computer Science > Computation and Language

arXiv:2110.01552 (cs)

[Submitted on 4 Oct 2021]

Title:Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA

Authors:Manuel R. Ciosici, Joe Cecil, Alex Hedges, Dong-Ho Lee, Marjorie Freedman, Ralph Weischedel

View PDF

Abstract:Our goal is to deliver a new task and leaderboard to stimulate research on question answering and pre-trained language models (PTLMs) to understand a significant instructional document, e.g., an introductory college textbook or a manual. PTLMs have shown great success in many question-answering tasks, given significant supervised training, but much less so in zero-shot settings. We propose a new task that includes two college-level introductory texts in the social sciences (American Government 2e) and humanities (U.S. History), hundreds of true/false statements based on review questions written by the textbook authors, validation/development tests based on the first eight chapters of the textbooks, blind tests based on the remaining textbook chapters, and baseline results given state-of-the-art PTLMs. Since the questions are balanced, random performance should be ~50%. T5, fine-tuned with BoolQ achieves the same performance, suggesting that the textbook's content is not pre-represented in the PTLM. Taking the exam closed book, but having read the textbook (i.e., adding the textbook to T5's pre-training), yields at best minor improvement (56%), suggesting that the PTLM may not have "understood" the textbook (or perhaps misunderstood the questions). Performance is better (~60%) when the exam is taken open-book (i.e., allowing the machine to automatically retrieve a paragraph and use it to answer the question).

Comments:	Identical to the EMNLP 2021 version
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2110.01552 [cs.CL]
	(or arXiv:2110.01552v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2110.01552

Submission history

From: Manuel Ciosici [view email]
[v1] Mon, 4 Oct 2021 16:45:28 UTC (41 KB)

Computer Science > Computation and Language

Title:Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Perhaps PTLMs Should Go to School -- A Task to Assess Open Book and Closed Book QA

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators