Multimodal Prototypical Networks for Few-shot Learning

11/17/2020
by   Frederik Pahde, et al.
0

Although providing exceptional results for many computer vision tasks, state-of-the-art deep learning algorithms catastrophically struggle in low data scenarios. However, if data in additional modalities exist (e.g. text) this can compensate for the lack of data and improve the classification results. To overcome this data scarcity, we design a cross-modal feature generation framework capable of enriching the low populated embedding space in few-shot scenarios, leveraging data from the auxiliary modality. Specifically, we train a generative model that maps text data into the visual feature space to obtain more reliable prototypes. This allows to exploit data from additional modalities (e.g. text) during training while the ultimate task at test time remains classification with exclusively visual data. We show that in such cases nearest neighbor classification is a viable approach and outperform state-of-the-art single-modal and multimodal few-shot learning methods on the CUB-200 and Oxford-102 datasets.

READ FULL TEXT
research
11/22/2018

Self Paced Adversarial Training for Multimodal Few-shot Learning

State-of-the-art deep learning algorithms yield remarkable results in ma...
research
01/16/2023

Multimodality Helps Unimodality: Cross-Modal Few-Shot Learning with Multimodal Models

The ability to quickly learn a new task with minimal instruction - known...
research
06/13/2018

Cross-modal Hallucination for Few-shot Fine-grained Recognition

State-of-the-art deep learning algorithms generally require large amount...
research
04/06/2019

Few-Shot Learning via Saliency-guided Hallucination of Samples

Learning new concepts from a few of samples is a standard challenge in c...
research
12/29/2022

Learning Multimodal Data Augmentation in Feature Space

The ability to jointly learn from multiple modalities, such as text, aud...
research
02/19/2019

Adaptive Cross-Modal Few-Shot Learning

Metric-based meta-learning techniques have successfully been applied to ...
research
06/15/2022

Rethinking Generalization in Few-Shot Classification

Single image-level annotations only correctly describe an often small su...

Please sign up or login with your details

Forgot password? Click here to reset