From:Internet Info Agency 2026-05-13 17:34:00
On May 13, Xiaomi officially launched and open-sourced XiaomiOneVL, a one-step latent-space language-vision reasoning framework. This framework unifies multiple technical approaches—including Vision-Language-Action (VLA), world models, and latent-space reasoning—within a single architecture for the first time, achieving performance improvements in perception, reasoning, and planning tasks for autonomous driving. XiaomiOneVL attains state-of-the-art (SOTA) results on three major benchmarks: ROADWork, Impromptu, and Alpamayo-R1, and demonstrates strong performance on the NAVSIM benchmark. Its reasoning accuracy surpasses explicit Chain-of-Thought (CoT) methods, while its inference speed matches that of latent-space CoT approaches that predict answers directly without intermediate reasoning steps. The framework supports dual interpretability in both language and vision, enabling it to simultaneously explain decision rationales in text and visualize future scenarios through predicted images. Xiaomi has open-sourced the model weights, training and inference code for XiaomiOneVL, along with its technical report and project homepage, making them available to the broader research and industry community.

Zeekr 7X Overheats at Ningbo Charging Station; Involved Vehicle Had Unrepaired Collision Damage
Mercedes-AMG Unveils Teaser of New All-Electric SUV with Triple-Motor System
Mitsubishi’s ASX VR-e EV, Based on Foxtron Bria, Heads to Australia
Beijing Hyundai Ioniq V Opens Pre-sales on August 21, Launches in September
BYD Seal 06 2027 Model Launches August 11, Starting at RMB 105,000
XPeng G9L Unveils Five New Colors, Reveals Dimensions and Cabin Features
China's Renewable Hydrogen Production Exceeds 1.4 Million Tons, with 620 Refueling Stations