research
∙
11/25/2022
XKD: Cross-modal Knowledge Distillation with Domain Alignment for Video Representation Learning
We present XKD, a novel self-supervised framework to learn meaningful re...
research
∙
11/09/2021
Self-Supervised Audio-Visual Representation Learning with Relaxed Cross-Modal Temporal Synchronicity
We present CrissCross, a self-supervised framework for learning audio-vi...
research
∙
02/04/2020