Meta-learning for compressed language model: A multiple choice question answering study

Ming Yan,Yi Pan

doi:10.1016/j.neucom.2021.01.148

Abstract

Model compression is a promising approach for reducing the model size of pretrained-language-models (PLMs) on low resource edge devices and applications. Unfortunately, the compression process always accompanies a cost of performance degradation, especially for the low resource downstream tasks, i.e., multiple-choice question answering. To address the degradation issue of model compression on PLMs, we proposed an end-to-end reptile (ETER) meta-learning approach to improving the performance of PLMs on the low resource multiple-choice question answering task. Specifically, our ETER improves the traditional two-stage meta-learning to an end-to-end manner, integrating the target finetuning stage into the meta training stage. To strengthen the generic meta-learning, ETER employs two-level meta-task construction from instance-level and domain-level to enrich its task generalization. What is more, ETER optimizes meta-learning by parameter constraints to reduce its parameter learning space. Experiments demonstrate that ETER significantly improved the performance of compressed PLMs and achieved large superiority over the baselines on different datasets.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Neurocomputing	Publication Date: Nov 4, 2021
Citations: 3	License type: publisher-specific-oa

R Discovery Prime

R Discovery Prime

Meta-learning for compressed language model: A multiple choice question answering study

Abstract

Talk to us

Similar Papers

More From: Neurocomputing

Lead the way for us

Similar Papers

Improving Subject-Area Question Answering with External Knowledge
Xiaoman Pan ... Dong Yu
-
Xiaoman Pan, et. al.Xiaoman Pan ... Dong Yu
01 Jan 2019
01 Jan 2019

Enhancing Multiple-Choice Question Answering with Causal Knowledge
Dhairya Dalal ... Paul Buitelaar
-
Dhairya Dalal, et. al.Dhairya Dalal ... Paul Buitelaar
01 Jan 2020
01 Jan 2020

A Multi-tasking and Multi-stage Chinese Minority Pre-trained Language Model
Bin Li ... Bin Sun
-
Bin Li, et. al.Bin Li ... Bin Sun
01 Jan 2021
01 Jan 2021

Training Compact Models for Low Resource Entity Tagging using Pre-trained Language Models
Peter Izsak ... Moshe Wasserblat
-
Peter Izsak, et. al.Peter Izsak ... Moshe Wasserblat
01 Dec 2019
01 Dec 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Meta-learning for compressed language model: A multiple choice question answering study

Abstract

Talk to us

Similar Papers

More From: Neurocomputing