ST-MoE is an AI model developed by Google, Google Brain and Google Research (United States), first published in February 2022. It works in the language domain, on tasks such as language modeling/generation.
Training it took an estimated 2.9×10²³ FLOP of compute (estimation method: operation counting). The model has 269,000,000,000 parameters. It was trained on roughly 1.5T datapoints.
Access: Unreleased. Its weights are not openly released. The reference paper has 375 citations. Epoch AI rates the confidence of this record as likely.