GLM-130B is an AI model developed by Tsinghua University (China), first published in August 2022. It works in the language domain, on tasks such as language modeling/generation and translation.
Training it took an estimated 3.5×10²³ FLOP of compute (estimation method: operation counting,hardware). The model has 130,000,000,000 parameters. It was trained on roughly 152B datapoints. Training ran on 768 NVIDIA A100 SXM4 40 GB for about 1.4K hours. The compute alone is estimated at $820K in 2023 dollars.
Access: Open weights (non-commercial). Its weights are openly available. The reference paper has 1,264 citations. Epoch AI rates the confidence of this record as confident.