On Leveraging Variational Graph Embeddings for Open World Compositional Zero-Shot Learning

04/23/2022
by   Muhammad Umer Anwaar, et al.
25

Humans are able to identify and categorize novel compositions of known concepts. The task in Compositional Zero-Shot learning (CZSL) is to learn composition of primitive concepts, i.e. objects and states, in such a way that even their novel compositions can be zero-shot classified. In this work, we do not assume any prior knowledge on the feasibility of novel compositions i.e.open-world setting, where infeasible compositions dominate the search space. We propose a Compositional Variational Graph Autoencoder (CVGAE) approach for learning the variational embeddings of the primitive concepts (nodes) as well as feasibility of their compositions (via edges). Such modelling makes CVGAE scalable to real-world application scenarios. This is in contrast to SOTA method, CGE, which is computationally very expensive. e.g.for benchmark C-GQA dataset, CGE requires 3.94 x 10^5 nodes, whereas CVGAE requires only 1323 nodes. We learn a mapping of the graph and image embeddings onto a common embedding space. CVGAE adopts a deep metric learning approach and learns a similarity metric in this space via bi-directional contrastive loss between projected graph and image embeddings. We validate the effectiveness of our approach on three benchmark datasets.We also demonstrate via an image retrieval task that the representations learnt by CVGAE are better suited for compositional generalization.

READ FULL TEXT

page 1

page 4

page 7

page 8

research
01/29/2021

Open World Compositional Zero-Shot Learning

Compositional Zero-Shot learning (CZSL) requires to recognize state-obje...
research
05/03/2021

Learning Graph Embeddings for Open World Compositional Zero-Shot Learning

Compositional Zero-Shot learning (CZSL) aims to recognize unseen composi...
research
02/03/2021

Learning Graph Embeddings for Compositional Zero-shot Learning

In compositional zero-shot learning, the goal is to recognize unseen com...
research
11/23/2022

InDiReCT: Language-Guided Zero-Shot Deep Metric Learning for Images

Common Deep Metric Learning (DML) datasets specify only one notion of si...
research
06/01/2021

Independent Prototype Propagation for Zero-Shot Compositionality

Humans are good at compositional zero-shot reasoning; someone who has ne...
research
11/05/2022

Simple Primitives with Feasibility- and Contextuality-Dependence for Open-World Compositional Zero-shot Learning

The task of Compositional Zero-Shot Learning (CZSL) is to recognize imag...
research
03/31/2022

Do Vision-Language Pretrained Models Learn Primitive Concepts?

Vision-language pretrained models have achieved impressive performance o...

Please sign up or login with your details

Forgot password? Click here to reset