Transfer Learning for Sequences via Learning to Collocate

02/25/2019
by   Wanyun Cui, et al.
0

Transfer learning aims to solve the data sparsity for a target domain by applying information of the source domain. Given a sequence (e.g. a natural language sentence), the transfer learning, usually enabled by recurrent neural network (RNN), represents the sequential information transfer. RNN uses a chain of repeating cells to model the sequence data. However, previous studies of neural network based transfer learning simply represents the whole sentence by a single vector, which is unfeasible for seq2seq and sequence labeling. Meanwhile, such layer-wise transfer learning mechanisms lose the fine-grained cell-level information from the source domain. In this paper, we proposed the aligned recurrent transfer, ART, to achieve cell-level information transfer. ART is under the pre-training framework. Each cell attentively accepts transferred information from a set of positions in the source domain. Therefore, ART learns the cross-domain word collocations in a more flexible way. We conducted extensive experiments on both sequence labeling tasks (POS tagging, NER) and sentence classification (sentiment analysis). ART outperforms the state-of-the-arts over all experiments.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/23/2018

DT-LET: Deep Transfer Learning by Exploring where to Transfer

Previous transfer learning methods based on deep network assume the know...
research
09/10/2019

Fine-grained Knowledge Fusion for Sequence Labeling Domain Adaptation

In sequence labeling, previous domain adaptation methods focus on the ad...
research
02/14/2019

Transfer Learning for Sequence Labeling Using Source Model and Target Data

In this paper, we propose an approach for transferring the knowledge of ...
research
10/09/2018

An Instance Transfer based Approach Using Enhanced Recurrent Neural Network for Domain Named Entity Recognition

Recently, neural networks have shown promising results for named entity ...
research
05/09/2022

Transfer Learning Based Efficient Traffic Prediction with Limited Training Data

Efficient prediction of internet traffic is an essential part of Self Or...
research
03/19/2016

How Transferable are Neural Networks in NLP Applications?

Transfer learning is aimed to make use of valuable knowledge in a source...
research
02/23/2021

A Novel Deep Learning Method for Textual Sentiment Analysis

Sentiment analysis is known as one of the most crucial tasks in the fiel...

Please sign up or login with your details

Forgot password? Click here to reset