Heuristic Reinforcement Learning is an AI model developed by Purdue University (United States), first published in October 1965. It works in the robotics domain, on tasks such as system control.
Training it took an estimated 1.1×10⁶ FLOP of compute (estimation method: hardware).
Epoch AI rates the confidence of this record as speculative.