Physical Intelligence (π) has announced the development of π0 (pi-zero), a new general-purpose robot foundation model. Unlike traditional robotic software designed for specific tasks, π0 is a Vision-Language-Action (VLA) model capable of controlling various robot configurations, from dexterous hands to mobile manipulators, across a wide range of physical tasks.
A Foundation for General Robotics
The π0 model is trained on the largest and most diverse robot dataset to date, incorporating data from numerous robot platforms and environments. This allows the model to understand complex instructions and translate visual inputs directly into physical actions. By leveraging large-scale pre-training similar to LLMs, π0 demonstrates an unprecedented ability to generalize across different hardware and novel scenarios.
Speed and Efficiency with π0-FAST
Alongside the primary model, the team introduced π0-FAST, a high-frequency version optimized for real-time applications. While π0 handles complex reasoning and long-horizon planning, π0-FAST ensures the low-latency response times required for fluid, reactive robotic movement. This dual approach aims to bridge the gap between high-level intelligence and reliable physical execution.
The introduction of these models represents a significant step toward a "universal brain" for robots, potentially reducing the need for specialized programming for every new robotic application.








