Webly Supervised Semantic Embeddings for Large Scale Zero-Shot Learning

08/06/2020
by   Yannick Le Cacheux, et al.
8

Zero-shot learning (ZSL) makes object recognition in images possible in absence of visual training data for a part of the classes from a dataset. When the number of classes is large, classes are usually represented by semantic class prototypes learned automatically from unannotated text collections. This typically leads to much lower performances than with manually designed semantic prototypes such as attributes. While most ZSL works focus on the visual aspect and reuse standard semantic prototypes learned from generic text collections, we focus on the problem of semantic class prototype design for large scale ZSL. More specifically, we investigate the use of noisy textual metadata associated to photos as text collections, as we hypothesize they are likely to provide more plausible semantic embeddings for visual classes if exploited appropriately. We thus make use of a source-based voting strategy to improve the robustness of semantic prototypes. Evaluation on the large scale ImageNet dataset shows a significant improvement in ZSL performances over two strong baselines, and over usual semantic embeddings used in previous works. We show that this improvement is obtained for several embedding methods, leading to state of the art results when one uses automatically created visual and text features.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/03/2022

Semantically Grounded Visual Embeddings for Zero-Shot Learning

Zero-shot learning methods rely on fixed visual and semantic embeddings,...
research
10/06/2020

Using Sentences as Semantic Representations in Large Scale Zero-Shot Learning

Zero-shot learning aims to recognize instances of unseen classes, for wh...
research
04/05/2016

Less is more: zero-shot learning from online textual documents with noise suppression

Classifying a visual concept merely from its associated online textual s...
research
04/02/2018

Hierarchical Novelty Detection for Visual Object Recognition

Deep neural networks have achieved impressive success in large-scale vis...
research
08/09/2017

An evaluation of large-scale methods for image instance and class discovery

This paper aims at discovering meaningful subsets of related images from...
research
04/21/2021

Revisiting Document Representations for Large-Scale Zero-Shot Learning

Zero-shot learning aims to recognize unseen objects using their semantic...
research
04/05/2023

What's in a Name? Beyond Class Indices for Image Recognition

Existing machine learning models demonstrate excellent performance in im...

Please sign up or login with your details

Forgot password? Click here to reset