S4 is an AI model developed by Stanford University (United States), first published in October 2021. It works in the language domain, on tasks such as language modeling/generation.
Training it took an estimated 7.8×10¹⁹ FLOP of compute. The model has 249,000,000 parameters. It was trained on roughly 103M datapoints. Training ran on 8 NVIDIA A100.
Access: Open weights (unrestricted). Its weights are openly available. The reference paper has 3,576 citations. Epoch AI rates the confidence of this record as likely.