End-to-end data-dependent routing in multi-path neural networks

Dumindu Tissera,Subha Fernando,Alex Xavier,Kasun Vithanage,Ranga Rodrigo,Rukshan Wijesinghe

doi:10.1007/s00521-023-08381-8

Abstract

Neural networks are known to give better performance with increased depth due to their ability to learn more abstract features. Although the deepening of networks has been well established, there is still room for efficient feature extraction within a layer, which would reduce the need for mere parameter increment. The conventional widening of networks by having more filters in each layer introduces a quadratic increment of parameters. Having multiple parallel convolutional/dense operations in each layer solves this problem, but without any context-dependent allocation of input among these operations: The parallel computations tend to learn similar features making the widening process less effective. Therefore, we propose the use of multi-path neural networks with data-dependent resource allocation from parallel computations within layers, which also lets an input be routed end-to-end through these parallel paths. To do this, we first introduce a cross-prediction-based algorithm between parallel tensors of subsequent layers. Second, we further reduce the routing overhead by introducing feature-dependent cross-connections between parallel tensors of successive layers. Using image recognition tasks, we show that our multi-path networks show superior performance to existing widening and adaptive feature extraction, even ensembles and deeper networks at similar complexity.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

End-to-end data-dependent routing in multi-path neural networks

Abstract

Talk to us

Similar Papers

More From: Neural Computing and Applications

Lead the way for us

Similar Papers

An On-Line Unsupervised Neural Network to Adaptive Feature Extraction of Lower SNR DS-SS signals
Tianqi Zhang ... Qianbin Chen
-
Tianqi Zhang, et. al.Tianqi Zhang ... Qianbin Chen
01 Oct 2006
01 Oct 2006

Adaptive Feature Extraction of Lower SNR DS-CDMA signals
Tianqi Zhang ... Zhengzhong Zhou
-
Tianqi Zhang, et. al.Tianqi Zhang ... Zhengzhong Zhou
01 Jan 2006
01 Jan 2006

A Neural Network Method to Adaptive Feature Extraction of Weak DS-CDMA Signals
Tianqi Zhang ... Xuesong Li
-
Tianqi Zhang, et. al.Tianqi Zhang ... Xuesong Li
01 Jan 2008
01 Jan 2008

A fault diagnosis method for nuclear power plant rotating machinery based on adaptive deep feature extraction and multiple support vector machines
Wenzhe Yin ... Miyombo Ernest Miyombo
Progress in Nuclear Energy | VOL. 164
Wenzhe Yin, et. al.Wenzhe Yin ... Miyombo Ernest Miyombo
09 Sep 2023
Progress in Nuclear Energy | VOL. 164

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

End-to-end data-dependent routing in multi-path neural networks

Abstract

Talk to us

Similar Papers

More From: Neural Computing and Applications