FLAN 137B is an AI model developed by Google Research (United States), first published in September 2021. It works in the language domain, on tasks such as language modeling, question answering and language modeling/generation. It counts among the frontier models: the systems trained with the most compute of their moment.
Training it took an estimated 2×10²⁴ FLOP of compute (estimation method: operation counting). The model has 137,000,000,000 parameters. It was trained on roughly 2.5T datapoints. Training ran on Google TPU v3.
Access: Unreleased. Its weights are not openly released. It is built on top of LaMDA. The reference paper has 5,011 citations. Epoch AI rates the confidence of this record as confident.