He Xiaopeng told reporters at Auto China 2026 on April 25 that Xpeng’s Vision Language Action system already outperforms Tesla’s Full Self-Driving in complex Chinese driving scenarios , and staked a public wager on closing the remaining gap entirely by August 30. The bet is not metaphorical. He set it in December 2025 after personally testing Tesla’s FSD V14.2 in Silicon Valley and returning to China convinced that the technology had reached “near-Level 4” performance. The terms are public and specific: if Xpeng’s VLA 2.0 system achieves within China the same overall road experience that Tesla’s FSD V14.2 delivers in Silicon Valley by August 30, 2026, He will build a Chinese-style canteen in Silicon Valley modeled after the cafeteria at Xpeng headquarters. If VLA falls short, Liu Xianming, Xpeng’s head of intelligent driving, will run naked across the Golden Gate Bridge. He invited Elon Musk to come test the result personally. As wagers go, it communicates the timeline with unusual clarity. Xpeng launched VLA 2.0 in March 2026 and showcased it at Auto China 2026, the Beijing International Automotive Exhibition running through May 3. The Vision Language Action architecture replaces the fragmented pipeline approach , where separate AI modules handle perception, planning, and action in sequence , with a single end-to-end model that processes visual input and generates vehicle actions the way a human driver does: continuously, in context, without handoffs between subsystems. The practical result, in Xpeng’s own comparison data, is measurable: on a 20-kilometer complex urban test route, Tesla FSD V13.2.9 required five driver takeovers while Xpeng’s second-generation VLA required one. Morgan Stanley analysts who tested the system in early March 2026 described themselves as surprised by how quickly the gap with Tesla had closed. He Xiaopeng has publicly welcomed all competitors, including Musk, to experience the system