In-Context Imitation Learning via Next-Token Prediction

Fu, Letian; Huang, Huang; Datta, Gaurav; Chen, Lawrence Yunliang; Panitch, William Chung-Ho; Liu, Fangchen; Li, Hui; Goldberg, Ken

Computer Science > Robotics

arXiv:2408.15980 (cs)

[Submitted on 28 Aug 2024 (v1), last revised 27 Sep 2024 (this version, v2)]

Title:In-Context Imitation Learning via Next-Token Prediction

Authors:Letian Fu, Huang Huang, Gaurav Datta, Lawrence Yunliang Chen, William Chung-Ho Panitch, Fangchen Liu, Hui Li, Ken Goldberg

View PDF HTML (experimental)

Abstract:We explore how to enhance next-token prediction models to perform in-context imitation learning on a real robot, where the robot executes new tasks by interpreting contextual information provided during the input phase, without updating its underlying policy parameters. We propose In-Context Robot Transformer (ICRT), a causal transformer that performs autoregressive prediction on sensorimotor trajectories without relying on any linguistic data or reward function. This formulation enables flexible and training-free execution of new tasks at test time, achieved by prompting the model with sensorimotor trajectories of the new task composing of image observations, actions and states tuples, collected through human teleoperation. Experiments with a Franka Emika robot demonstrate that the ICRT can adapt to new tasks specified by prompts, even in environment configurations that differ from both the prompt and the training data. In a multitask environment setup, ICRT significantly outperforms current state-of-the-art next-token prediction models in robotics on generalizing to unseen tasks. Code, checkpoints and data are available on this https URL

Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2408.15980 [cs.RO]
	(or arXiv:2408.15980v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2408.15980

Submission history

From: Letian Fu [view email]
[v1] Wed, 28 Aug 2024 17:50:19 UTC (10,252 KB)
[v2] Fri, 27 Sep 2024 20:10:08 UTC (10,343 KB)

Computer Science > Robotics

Title:In-Context Imitation Learning via Next-Token Prediction

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:In-Context Imitation Learning via Next-Token Prediction

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators