EPISODE · Jul 20, 2026 · 20 MIN
Xiaomi-Robotics-1: A Scalable Vision-Language-Action Foundation Model for Mobile Manipulation
from Embodied AI 101 · host Shaoqing Tan
A scalable vision-language-action (VLA) foundation model pretrained on over 100k hours of real-world manipulation trajectories, enabling out-of-the-box mobile manipulation capabilities.
Embed this episode
NOW PLAYING
Xiaomi-Robotics-1: A Scalable Vision-Language-Action Foundation Model for Mobile Manipulation
0:00
20:06
1×
No transcript for this episode yet
Similar Episodes
No similar episodes found.
Similar Podcasts
Frequently Asked Questions
How long is this episode of Embodied AI 101?
This episode is 20 minutes long.
When was this Embodied AI 101 episode published?
This episode was published on July 20, 2026.
Can I download this Embodied AI 101 episode?
Yes. Use the download control on the episode player to save the publisher-provided media file.
URL copied to clipboard!