LLaVA 1.5 is an AI model developed by University of Wisconsin Madison and Microsoft Research (United States), first published in November 2023. It works in the multimodal, language and vision domain, on tasks such as chat, question answering and visual question answering.
Training it took an estimated 7.8×10²² FLOP of compute (estimation method: hardware). The model has 13,000,000,000 parameters. Training ran on 8 NVIDIA A100 for about 24 hours.
Access: Open weights (restricted use). Its weights are openly available. It is built on top of Vicuna-13B v0. The reference paper has 5,077 citations. Epoch AI rates the confidence of this record as confident.