tools
π0.5
π0.5 is a vision-language-action model for controlling mobile manipulators in unfamiliar homes and other environments.

π0.5 combines language, visual, and robotic-action data to control robots at both high and low levels. It can interpret broad instructions such as cleaning a bedroom, break them into subtasks, and produce motor commands.
Physical Intelligence evaluated it on tasks including putting dishes away, making beds, cleaning floors, rearranging objects, and wiping spills in homes absent from its training data. The model is not presented as a standalone chatbot or hosted consumer service; robot hardware, integration, deployment, and access arrangements are not listed.
Features
- Controls mobile manipulators in homes not seen during training
- Accepts high-level and detailed natural-language commands
- Breaks long tasks into semantic subtasks
- Outputs continuous low-level robot joint actions
- Uses a flow-matching action expert for motor control
- Co-trains on web, multimodal, and cross-robot data
- Handles object rearrangement and cleaning tasks
- Supports high- and low-level control in one VLA model
Use cases
- Clean kitchens and bedrooms in unfamiliar homes
- Put dishes and other objects in their intended places
- Make a bed from a high-level instruction
- Move specified objects into drawers
- Wipe spills with a sponge
- Adapt robot behavior to new environments