Transformer-XL Large + Phrase Induction is an AI model developed by Massachusetts Institute of Technology (MIT) and University of Illinois Urbana-Champaign (UIUC) (United States), first published in June 2019. It works in the language domain, on tasks such as language modeling/generation.
Training it took an estimated 3.8×10²⁰ FLOP of compute (estimation method: operation counting). The model has 257,000,000 parameters. It was trained on roughly 103M datapoints.
Access: Unreleased. Its weights are not openly released. It is built on top of Transformer-XL (257M). The reference paper has 14 citations. Epoch AI rates the confidence of this record as speculative.