LLaVA-NeXT-34B (LLaVA-1.6) is an AI model developed by University of Wisconsin Madison, ByteDance, Nanyang Technological University and University of California (UC) Berkeley (United States, China and Singapore), first published in January 2024. It works in the multimodal, language and vision domain, on tasks such as visual question answering, chat and question answering.
Training it took an estimated 2.6×10²⁰ FLOP of compute (estimation method: hardware). The model has 34,750,000,000 parameters. It was trained on roughly 89.3M datapoints. Training ran on 32 NVIDIA A100 for about 24 hours.
Access: Open weights (unrestricted). Its weights are openly available. Epoch AI rates the confidence of this record as likely.