Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance

Wang, Zhe; Wang, Haozhu; Qi, Yanjun

Computer Science > Machine Learning

arXiv:2412.00979 (cs)

[Submitted on 1 Dec 2024 (v1), last revised 13 Dec 2024 (this version, v2)]

Title:Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance

Authors:Zhe Wang, Haozhu Wang, Yanjun Qi

View PDF HTML (experimental)

Abstract:Decision transformers recast reinforcement learning as a conditional sequence generation problem, offering a simple but effective alternative to traditional value or policy-based methods. A recent key development in this area is the integration of prompting in decision transformers to facilitate few-shot policy generalization. However, current methods mainly use static prompt segments to guide rollouts, limiting their ability to provide context-specific guidance. Addressing this, we introduce a hierarchical prompting approach enabled by retrieval augmentation. Our method learns two layers of soft tokens as guiding prompts: (1) global tokens encapsulating task-level information about trajectories, and (2) adaptive tokens that deliver focused, timestep-specific instructions. The adaptive tokens are dynamically retrieved from a curated set of demonstration segments, ensuring context-aware guidance. Experiments across seven benchmark tasks in the MuJoCo and MetaWorld environments demonstrate the proposed approach consistently outperforms all baseline methods, suggesting that hierarchical prompting for decision transformers is an effective strategy to enable few-shot policy generalization.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2412.00979 [cs.LG]
	(or arXiv:2412.00979v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2412.00979

Submission history

From: Zhe Wang [view email]
[v1] Sun, 1 Dec 2024 22:02:07 UTC (363 KB)
[v2] Fri, 13 Dec 2024 03:31:51 UTC (363 KB)

Computer Science > Machine Learning

Title:Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators