OpenVLA is an AI model developed by Stanford University, University of California (UC) Berkeley, Toyota Research Institute, Google DeepMind, Massachusetts Institute of Technology (MIT) and Physical Intelligence (United States), first published in June 2024. It works in the robotics, vision and language domain, on tasks such as robotic manipulation.
Training it took an estimated 1.1×10²³ FLOP of compute (estimation method: hardware). The model has 7,188,100,000 parameters. Training ran on 64 NVIDIA A100 for about 336 hours.
Access: Open weights (unrestricted). Its weights are openly available. It is built on top of Llama 2-7B. The reference paper has 2,123 citations. Epoch AI rates the confidence of this record as confident.