Drastically Reducing the Number of Trainable Parameters in Deep CNNs by   Inter-layer Kernel-sharing

Alireza Azadbakht; Saeed Reza Kheradpisheh; Ismail Khalfaoui-Hassani,; Timoth\'ee Masquelier

arXiv:2210.14151·cs.CV·October 26, 2022

Drastically Reducing the Number of Trainable Parameters in Deep CNNs by Inter-layer Kernel-sharing

Alireza Azadbakht, Saeed Reza Kheradpisheh, Ismail Khalfaoui-Hassani,, Timoth\'ee Masquelier

PDF

Open Access 1 Repo

TL;DR

This paper proposes a simple kernel-sharing method between isomorphic layers in deep CNNs to drastically reduce trainable parameters, enabling efficient edge computing with minimal accuracy loss.

Contribution

Introducing kernel-sharing between isomorphic layers in CNNs as a novel way to reduce parameters and improve regularization.

Findings

01

Significant parameter reduction with minimal accuracy loss

02

Effective regularization reducing overfitting

03

Suitable for memory-constrained edge devices

Abstract

Deep convolutional neural networks (DCNNs) have become the state-of-the-art (SOTA) approach for many computer vision tasks: image classification, object detection, semantic segmentation, etc. However, most SOTA networks are too large for edge computing. Here, we suggest a simple way to reduce the number of trainable parameters and thus the memory footprint: sharing kernels between multiple convolutional layers. Kernel-sharing is only possible between ``isomorphic" layers, i.e.layers having the same kernel size, input and output channels. This is typically the case inside each stage of a DCNN. Our experiments on CIFAR-10 and CIFAR-100, using the ConvMixer and SE-ResNet architectures show that the number of parameters of these models can drastically be reduced with minimal cost on accuracy. The resulting networks are appealing for certain edge computing applications that are subject to…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

alirezaazadbakht/kernel-sharing
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Neural Network Applications · Domain Adaptation and Few-Shot Learning · Advanced Image and Video Retrieval Techniques

MethodsDiffusion-Convolutional Neural Networks