Technology term
Vision-Language-Action Model (VLA)
A vision-language-action model is an artificial-intelligence system that combines visual perception, language understanding, and action generation so a robot can interpret a scene and an instruction, then produce physical control commands.
English
Vision-Language-Action Model
Arabic
نموذج الرؤية واللغة والحركة
First appeared in
Robots Pause to Think. This New AI Method Lets Them Plan While Moving