系列:VLA Decoding Notes

VLA Decoding Notes (3): The π Family — Physical Intelligence’s VLA Lineage

1. π0: A “Universal Interface” for VLA (2024-10)

Physical Intelligence’s π0 set the baseline:

2. π0-Fast: Discrete Speedup

Adds action discretization + autoregressive on top of π0 for inference speed — an engineering trade-off between the continuous vs discrete routes, for latency-sensitive settings.

3. π0.5: Cross-Embodiment + Task Decomposition (2025-04)

4. π0.6*: RL Specialization (RECAP)

Uses reinforcement learning to specialize the trained VLA (RECAP-style), pushing specific-task success rates further. The typical second half of “general pre-train + specialist post-train”.

5. π0.7: Memory + Compositional Generalization (2026-04)

π0.7 is the current apex:

π family evolution π0FlowMatch 50Hz π0-Fastdiscretize π0.5cross-embod. π0.6*RL spec. π0.7memory+comp.

Fig: The π lineage — backbone upgrade, cross-embodiment, RL specialization, memory + compositional generalization

6. The Through-Line (One Line)

VLM backbone upgrade → cross-embodiment reuse → RL specialization → memory + compositional generalization. General VLA is moving from “can work” to “stronger than specialists”.

7. Bridge

The π family is the overseas benchmark. Next: domestic players — Ant Lingbo’s LingBot-VLA 2.0 “one brain, many machines”, Xiaomi’s open real-time VLA, Tencent HyVLA’s FlowPRO — plus a clarification of a common mix-up.

觉得有用?欢迎点赞、收藏,或请作者喝咖啡 ☕️

支付宝收款码

支付宝

微信收款码

微信

💬 留言

评论由 Giscus 驱动(基于 GitHub Discussions)。 当前仓库 NaphJohn/LLM-blog 尚未启用 Discussions:请在 GitHub 仓库 Settings → General → Features 勾选 Discussions 后刷新本页,评论区即自动显示。