CogVLM-17B is an AI model developed by Tsinghua University, Z.ai (Zhipu AI) and Beihang University (China), first published in November 2023. It works in the multimodal, vision and language domain, on tasks such as image captioning, visual question answering and chat.
Training it took an estimated 6.3×10²² FLOP of compute (estimation method: reported). The model has 17,000,000,000 parameters.
Access: Open weights (restricted use). Its weights are openly available. It is built on top of Vicuna-7B v0. The reference paper has 794 citations. Epoch AI rates the confidence of this record as confident.