Live
AI models

CogVLM-17B

Training compute
6.3×10²² FLOP
Parameters
17B
Published
Nov 6, 2023

CogVLM-17B is an AI model developed by Tsinghua University, Z.ai (Zhipu AI) and Beihang University (China), first published in November 2023. It works in the multimodal, vision and language domain, on tasks such as image captioning, visual question answering and chat.

Training it took an estimated 6.3×10²² FLOP of compute (estimation method: reported). The model has 17,000,000,000 parameters.

Access: Open weights (restricted use). Its weights are openly available. It is built on top of Vicuna-7B v0. The reference paper has 794 citations. Epoch AI rates the confidence of this record as confident.

Full record
Organization
Tsinghua University, Z.ai (Zhipu AI), Beihang University
Country of organization
China
Domain
Multimodal, Vision, Language
Task
Image captioning, Visual question answering, Chat
Training compute
6.3×10²² FLOP
Compute estimation method
Reported
Parameters
17,000,000,000
Numerical format
BF16
Model accessibility
Open weights (restricted use)
Open weights
Yes
Base model
Vicuna-7B v0
Citations
794
Epoch confidence
Confident
SourceEpoch AI, 'AI Models'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/data/ai-models. Licensed under CC BY 4.0.
← All ai models