Spatiotemporal fusion ConvNet is an AI model developed by Graz University of Technology and University of Oxford (Austria and United Kingdom), first published in June 2016. It works in the video domain, on tasks such as video and action recognition.
Epoch AI has no training-compute estimate for this model. It was trained on roughly 13K datapoints.
The reference paper has 2,538 citations.