MEAL: Multi-Model Ensemble via Adversarial Learning

Zhiqiang Shen,Xiangyang Xue,Zhankui He

doi:10.1609/aaai.v33i01.33014886

Abstract

Often the best performing deep neural models are ensembles of multiple base-level networks. Unfortunately, the space required to store these many networks, and the time required to execute them at test-time, prohibits their use in applications where test sets are large (e.g., ImageNet). In this paper, we present a method for compressing large, complex trained ensembles into a single network, where knowledge from a variety of trained deep neural networks (DNNs) is distilled and transferred to a single DNN. In order to distill diverse knowledge from different trained (teacher) models, we propose to use adversarial-based learning strategy where we define a block-wise training loss to guide and optimize the predefined student network to recover the knowledge in teacher models, and to promote the discriminator network to distinguish teacher vs. student features simultaneously. The proposed ensemble method (MEAL) of transferring distilled knowledge with adversarial learning exhibits three important advantages: (1) the student network that learns the distilled knowledge with discriminators is optimized better than the original model; (2) fast inference is realized by a single forward pass, while the performance is even better than traditional ensembles from multi-original models; (3) the student network can learn the distilled knowledge from a teacher model that has arbitrary structures. Extensive experiments on CIFAR-10/100, SVHN and ImageNet datasets demonstrate the effectiveness of our MEAL method. On ImageNet, our ResNet-50 based MEAL achieves top-1/5 21.79%/5.99% val error, which outperforms the original model by 2.06%/1.14%.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

MEAL: Multi-Model Ensemble via Adversarial Learning

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: Jul 17, 2019
Citations: 129

Similar Papers

FTBME: feature transferring based multi-model ensemble
A Yongquan Yang ... E Zhongxi Zheng
Multimedia Tools and Applications | VOL. 79
A Yongquan Yang, et. al.A Yongquan Yang ... E Zhongxi Zheng
12 Mar 2020
Multimedia Tools and Applications | VOL. 79

Adversarial learning: the impact of statistical sample selection techniques on neural ensembles
Shir Li Wang ... Hussein A Abbass
Evolving Systems | VOL. 1
Shir Li Wang, et. al.Shir Li Wang ... Hussein A Abbass
09 Aug 2010
Evolving Systems | VOL. 1

A Gift from Knowledge Distillation: Fast Optimization, Network Minimization and Transfer Learning
Junho Yim ... Donggyu Joo
-
Junho Yim, et. al.Junho Yim ... Donggyu Joo
01 Jul 2017
01 Jul 2017

Single Deterministic Neural Network with Hierarchical Gaussian Mixture Model for Uncertainty Quantification
Chunlin Ji ... Dingwei Gong
-
Chunlin Ji, et. al.Chunlin Ji ... Dingwei Gong
01 Jan 2021
01 Jan 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

MEAL: Multi-Model Ensemble via Adversarial Learning

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence