Skip to content

VLA model

A VLA (vision-language-action) model takes camera images and a language instruction and outputs robot actions directly.

  • VLA model
  • vision-language-action
  • robot foundation model
  • RT-2
  • pi-0
  • robotics AI
  • embodied AI
  • robot policy model