# VLA (Vision-Language-Action)

Foundation models that take camera input + a plain-English instruction (‘put the red parts in bin 3’) and output robot movements. VLA adoption tripled in a year — now in ~40% of new robot deployments. Programming is becoming prompting.

Canonical: https://robauto.ai/learn/iot-robotics/7

_IoT & Robotics: AI in the Physical World — lesson 7 of 20 (DEFINITION)_

Foundation models that take camera input + a plain-English instruction (‘put the red parts in bin 3’) and output robot movements. VLA adoption tripled in a year — now in ~40% of new robot deployments. Programming is becoming prompting.

Source: [State of Robotics 2026](https://www.roboticscenter.ai/state-of-robotics-2026?utm_source=robauto)

[Previous lesson](/learn/iot-robotics/6) · [Next lesson](/learn/iot-robotics/8) · [Course overview](/learn/iot-robotics) · [All courses](/learn)

---

(c) 2026 Robauto, Inc. — support@robauto.ai
Machine surfaces: https://robauto.ai/llms.txt · https://robauto.ai/llms-full.txt · https://robauto.ai/.well-known/api-catalog
