Back to cases
CASE 36 / General capabilities

π₀: from language instructions to laundry handling

A cross-embodiment vision-language-action model studies folding laundry, clearing tables and assembling boxes.

Physical Intelligence · π₀2024
π₀ paper, Figure 1: cross-embodiment manipulation tasks and model overview.
π₀ paper, Figure 1: cross-embodiment manipulation tasks and model overview.Physical Intelligence · π₀ paper, Figure 1
01 / STUDY

Mechanism and experiments

π₀ adds a flow-matching action component to a pretrained vision-language model and trains on data from several robots. The paper reports single-arm, bimanual and mobile-manipulation tasks. The openpi repository provides base models and selected task checkpoints, including an ALOHA towel-folding policy.

02 / CONTEXT

Test conditions and scope

Checkpoints target particular robots and tasks. Different cameras, grippers and control interfaces require adaptation and testing; the paper’s task coverage does not imply that all training data have been released.

03 / SOURCE

Papers and primary sources