Octo: An Open-Source Generalist Robot Policy

Octo Model Team; Ghosh, Dibya; Walke, Homer; Pertsch, Karl; Black, Kevin; Mees, Oier; Dasari, Sudeep; Hejna, Joey; Kreiman, Tobias; Xu, Charles; Luo, Jianlan; Tan, You Liang; Chen, Lawrence Yunliang; Sanketi, Pannag; Vuong, Quan; Xiao, Ted; Sadigh, Dorsa; Finn, Chelsea; Levine, Sergey

Computer Science > Robotics

arXiv:2405.12213 (cs)

[Submitted on 20 May 2024 (v1), last revised 26 May 2024 (this version, v2)]

Title:Octo: An Open-Source Generalist Robot Policy

Authors:Octo Model Team, Dibya Ghosh, Homer Walke, Karl Pertsch, Kevin Black, Oier Mees, Sudeep Dasari, Joey Hejna, Tobias Kreiman, Charles Xu, Jianlan Luo, You Liang Tan, Lawrence Yunliang Chen, Pannag Sanketi, Quan Vuong, Ted Xiao, Dorsa Sadigh, Chelsea Finn, Sergey Levine

View PDF HTML (experimental)

Abstract:Large policies pretrained on diverse robot datasets have the potential to transform robotic learning: instead of training new policies from scratch, such generalist robot policies may be finetuned with only a little in-domain data, yet generalize broadly. However, to be widely applicable across a range of robotic learning scenarios, environments, and tasks, such policies need to handle diverse sensors and action spaces, accommodate a variety of commonly used robotic platforms, and finetune readily and efficiently to new domains. In this work, we aim to lay the groundwork for developing open-source, widely applicable, generalist policies for robotic manipulation. As a first step, we introduce Octo, a large transformer-based policy trained on 800k trajectories from the Open X-Embodiment dataset, the largest robot manipulation dataset to date. It can be instructed via language commands or goal images and can be effectively finetuned to robot setups with new sensory inputs and action spaces within a few hours on standard consumer GPUs. In experiments across 9 robotic platforms, we demonstrate that Octo serves as a versatile policy initialization that can be effectively finetuned to new observation and action spaces. We also perform detailed ablations of design decisions for the Octo model, from architecture to training data, to guide future research on building generalist robot models.

Comments:	Project website: this https URL
Subjects:	Robotics (cs.RO); Machine Learning (cs.LG)
Cite as:	arXiv:2405.12213 [cs.RO]
	(or arXiv:2405.12213v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2405.12213

Submission history

From: Karl Pertsch [view email]
[v1] Mon, 20 May 2024 17:57:01 UTC (2,745 KB)
[v2] Sun, 26 May 2024 19:55:26 UTC (2,745 KB)

Computer Science > Robotics

Title:Octo: An Open-Source Generalist Robot Policy

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Octo: An Open-Source Generalist Robot Policy

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators