Aivora

#videomsn

VideoMSN

1 article

VideoMSN: Turning Image Classifiers into Efficient Video Learners via Super-Images
arXivAI Research

VideoMSN: Turning Image Classifiers into Efficient Video Learners via Super-Images

免 3D 結構與重構解碼器!VideoMSN 利用超大型圖片將影像分類器轉化為高效視訊學習器

VideoMSN is a self-supervised Masked Siamese Network that represents videos as grid-based 'super images', enabling standard image ViTs to learn powerful spatio-temporal video representations with up to 160x fewer pretraining epochs.

2 min read