Investigating Mysteries of CoT-Augmented Distillation

Wadhwa, Somin; Amir, Silvio; Wallace, Byron C.

Computer Science > Computation and Language

arXiv:2406.14511 (cs)

[Submitted on 20 Jun 2024 (v1), last revised 27 Sep 2024 (this version, v2)]

Title:Investigating Mysteries of CoT-Augmented Distillation

Authors:Somin Wadhwa, Silvio Amir, Byron C. Wallace

View PDF HTML (experimental)

Abstract:Eliciting "chain of thought" (CoT) rationales -- sequences of token that convey a "reasoning" process -- has been shown to consistently improve LLM performance on tasks like question answering. More recent efforts have shown that such rationales can also be used for model distillation: Including CoT sequences (elicited from a large "teacher" model) in addition to target labels when fine-tuning a small student model yields (often substantial) improvements. In this work we ask: Why and how does this additional training signal help in model distillation? We perform ablations to interrogate this, and report some potentially surprising results. Specifically: (1) Placing CoT sequences after labels (rather than before) realizes consistently better downstream performance -- this means that no student "reasoning" is necessary at test time to realize gains. (2) When rationales are appended in this way, they need not be coherent reasoning sequences to yield improvements; performance increases are robust to permutations of CoT tokens, for example. In fact, (3) a small number of key tokens are sufficient to achieve improvements equivalent to those observed when full rationales are used in model distillation.

Comments:	Accepted to EMNLP 2024
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2406.14511 [cs.CL]
	(or arXiv:2406.14511v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2406.14511

Submission history

From: Somin Wadhwa [view email]
[v1] Thu, 20 Jun 2024 17:15:46 UTC (15,599 KB)
[v2] Fri, 27 Sep 2024 20:13:16 UTC (15,587 KB)

Computer Science > Computation and Language

Title:Investigating Mysteries of CoT-Augmented Distillation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Investigating Mysteries of CoT-Augmented Distillation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators