资讯
小鹏刘先明:第二代 VLA 升级的核心,是让模型第一次建立起对时间维度的理解
📌 概要
<p>IT之家 8 月 27 日消息,小鹏物理 AI 分享暨第二代 VLA 全新版本体验日活动正在进行中,小鹏集团通用智能中心负责人刘先明登台发表演讲。他表示,物理 AI 真正的难题,是理解世界,并在充满限制的真实世界里行动。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/
<p>IT之家 8 月 27 日消息,小鹏物理 AI 分享暨第二代 VLA 全新版本体验日活动正在进行中,小鹏集团通用智能中心负责人刘先明登台发表演讲。他表示,物理 AI 真正的难题,是理解世界,并在充满限制的真实世界里行动。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/bce1c884-9a79-4c1c-bf94-1d04687b122a.jpg?x-bce-process=image/format,f_auto" /></p><p>刘先明称,小鹏在半年前提出物理 AI 能力公式:<strong>能力 = 模型 × 算力 × 数据 × 本体</strong>。其中,模型决定智能上限,数据提供世界经验,算力支撑训练与运行,本体让 AI 真正落地物理世界。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/66222fb5-8f92-49d0-97b0-316ecc923f85.jpg?x-bce-process=image/format,f_auto" /></p><p>刘先明认为,真实世界并不是一张张静止的图片,而是一个不断变化、连续演进的过程。模型想要像人类一样理解世界,必须首先理解什么是“时间”。<strong>第二代 VLA 升级的核心,是让模型第一次建立起对时间维度的理解</strong>。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/f82fac42-80d7-4223-a940-9d44118287d2.jpg?x-bce-process=image/format,f_auto" /></p><p>据介绍,小鹏物理世界基座模型,首次把“时间”纳入模型,AI 理解世界的维度,<strong>从“3D 空间”跃迁至“4D 时空”</strong>。时间越长、信息越多,模型就需要更快地反应、更高频地决策,并实现更强的安全能力;模型越大,泛化越强,计算需求也越高。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/c6a767e5-5d46-4dff-820d-8c8b14d07a58.jpg?x-bce-process=image/format,f_auto" /></p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/ed434e99-f7c0-4e80-9b18-a3612dbbbddd.jpg?x-bce-process=image/format,f_auto" /></p><p>IT之家注意到,以前模型看到的更多是当下的信息,<strong>第二代 VLA 全新引入 Infini-VLA 长时序架构,记得前 30 秒的世界</strong>,驾驶判断开始拥有更长的上下文,更多有效信息进入模型。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/3da3b2da-d8af-4ed2-a329-b6287cb964d9.jpg?x-bce-process=image/format,f_auto" /></p><p>道路情况是时刻在变化的,模型也需要边看边想,如果决策速度跟不上现实变化,也无法形成真正可靠的物理世界智能。所以,第二代 VLA 同步提升了模型推理效率,采用流式自回归推理,<strong>端到端的响应提速 300 %</strong>,实现“边看、边想、边行动”的并行处理。</p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/12bda841-d5f1-4433-96dc-186bcaedcbbc.jpg?x-bce-process=image/format,f_auto" /></p><p style="text-align: center;"><img class="no-alt-img" src="https://img.ithome.com/newsuploadfiles/2026/8/34b9ba52-b178-4557-b5b2-67bb66e46a6a.jpg?x-bce-process=image/format,f_auto" /></p>