Video Encoders

 Is Space-Time Attention All You Need for Video Understanding?

A Large-Scale Study on Unsupervised Spatiotemporal Representation Learning

Comments

Popular posts from this blog

Humans Learn From Task Descriptions and So Should Our Models

Divergence Frontiers for Evaluating Deep Generative Models