Tencent has announced the open-source release of three embodied foundation models during the WAIC 2026 event. These models include a Visual Language Model (VLM) designed for scene understanding, the RxBrain cognitive model that facilitates planning with visual states, and the VLA model, which supports continuous action at frequencies between 500 to 1000Hz.
This development is significant as it aims to enhance robot reaction speed and cognitive capabilities, addressing critical challenges in robotic performance. The introduction of these models is expected to advance the field of robotics by providing developers with powerful tools to improve the efficiency and effectiveness of robotic systems.
Looking ahead, the industry will be keen to observe how these open-source models are adopted and integrated into various robotic applications. No further timeline was disclosed at the time of publication.
Editor's Note
The open-sourcing of these foundational models by Tencent reflects a growing trend in the robotics industry towards collaborative development and innovation. By providing access to advanced cognitive and action models, Tencent is positioning itself as a key player in enhancing robotic capabilities, which could influence future technology adoption and competitive dynamics.
Leave a comment