Time Reversal as Self-Supervision

Nair, Suraj; Babaeizadeh, Mohammad; Finn, Chelsea; Levine, Sergey; Kumar, Vikash

Computer Science > Robotics

arXiv:1810.01128 (cs)

[Submitted on 2 Oct 2018 (v1), last revised 22 May 2020 (this version, v2)]

Title:Time Reversal as Self-Supervision

Authors:Suraj Nair, Mohammad Babaeizadeh, Chelsea Finn, Sergey Levine, Vikash Kumar

View PDF

Abstract:A longstanding challenge in robot learning for manipulation tasks has been the ability to generalize to varying initial conditions, diverse objects, and changing objectives. Learning based approaches have shown promise in producing robust policies, but require heavy supervision to efficiently learn precise control, especially from visual inputs. We propose a novel self-supervision technique that uses time-reversal to learn goals and provide a high level plan to reach them. In particular, we introduce the time-reversal model (TRM), a self-supervised model which explores outward from a set of goal states and learns to predict these trajectories in reverse. This provides a high level plan towards goals, allowing us to learn complex manipulation tasks with no demonstrations or exploration at test time. We test our method on the domain of assembly, specifically the mating of tetris-style block pairs. Using our method operating atop visual model predictive control, we are able to assemble tetris blocks on a physical robot using only uncalibrated RGB camera input, and generalize to unseen block pairs. this http URL

Comments:	7 pages, 10 figures
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI)
Cite as:	arXiv:1810.01128 [cs.RO]
	(or arXiv:1810.01128v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.1810.01128
Journal reference:	International Conference on Robotics and Automation, 2020

Submission history

From: Vikash Kumar [view email]
[v1] Tue, 2 Oct 2018 09:15:58 UTC (4,575 KB)
[v2] Fri, 22 May 2020 19:04:24 UTC (4,172 KB)

Computer Science > Robotics

Title:Time Reversal as Self-Supervision

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Time Reversal as Self-Supervision

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators