3 ms·
VideoCLIP: Contrastive Pre-Training for Zero-Shot Video-Text Understanding
- sharemywin 5y agoNot sure if this is the same thing? https://github.com/openai/CLIP https://github.com/openai/CLIP
- LuisMondragon 5y agoNot the same. CLIP is trained with pairs of images and texts, whereas VideoCLIP uses pairs of videos and texts.