LLaVA + LVIS-INSTRUCT4V is an AI model developed by Fudan University and University of Maryland (China and United States), first published in November 2023. It works in the multimodal, language and vision domain, on tasks such as language modeling/generation and visual question answering.
Epoch AI has no training-compute estimate for this model. The model has 13,000,000,000 parameters.
Access: Open weights (unrestricted). Its weights are openly available. It is built on top of LLaVA 1.5. The reference paper has 148 citations. Epoch AI rates the confidence of this record as likely.