MAGNeto: An Efficient Deep Learning Method for the Extractive Tags Summarization Problem

11/09/2020
by   Hieu Trong Phung, et al.
0

In this work, we study a new image annotation task named Extractive Tags Summarization (ETS). The goal is to extract important tags from the context lying in an image and its corresponding tags. We adjust some state-of-the-art deep learning models to utilize both visual and textual information. Our proposed solution consists of different widely used blocks like convolutional and self-attention layers, together with a novel idea of combining auxiliary loss functions and the gating mechanism to glue and elevate these fundamental components and form a unified architecture. Besides, we introduce a loss function that aims to reduce the imbalance of the training data and a simple but effective data augmentation technique dedicated to alleviates the effect of outliers on the final results. Last but not least, we explore an unsupervised pre-training strategy to further boost the performance of the model by making use of the abundant amount of available unlabeled data. Our model shows the good results as 90 50 Source code for reproducing the experiments is publicly available at: https://github.com/pixta-dev/labteam

READ FULL TEXT

page 6

page 9

page 15

research
04/14/2023

Toward Real-Time Image Annotation Using Marginalized Coupled Dictionary Learning

In most image retrieval systems, images include various high-level seman...
research
01/09/2022

Invariance encoding in sliced-Wasserstein space for image classification with limited training data

Deep convolutional neural networks (CNNs) are broadly considered to be s...
research
05/05/2021

DeepSMOTE: Fusing Deep Learning and SMOTE for Imbalanced Data

Despite over two decades of progress, imbalanced data is still considere...
research
09/01/2023

A Text-based Approach For Link Prediction on Wikipedia Articles

This paper present our work in the DSAA 2023 Challenge about Link Predic...
research
07/19/2021

Compound Figure Separation of Biomedical Images with Side Loss

Unsupervised learning algorithms (e.g., self-supervised learning, auto-e...
research
06/20/2022

Technical Report: Combining knowledge from Transfer Learning during training and Wide Resnets

In this report, we combine the idea of Wide ResNets and transfer learnin...
research
08/24/2022

Identifying Films with Noir Characteristics Using Audience's Tags on MovieLens

We consider the noir classification problem by exploring noir attributes...

Please sign up or login with your details

Forgot password? Click here to reset