3D Segmentation with Fully Trainable Gabor Kernels and Pearson's Correlation Coefficient

Wong, Ken C. L.; Moradi, Mehdi

doi:10.1007/978-3-031-21014-3_6

Electrical Engineering and Systems Science > Image and Video Processing

arXiv:2201.03644 (eess)

[Submitted on 10 Jan 2022 (v1), last revised 15 Dec 2022 (this version, v2)]

Title:3D Segmentation with Fully Trainable Gabor Kernels and Pearson's Correlation Coefficient

Authors:Ken C. L. Wong, Mehdi Moradi

View PDF

Abstract:The convolutional layer and loss function are two fundamental components in deep learning. Because of the success of conventional deep learning kernels, the less versatile Gabor kernels become less popular despite the fact that they can provide abundant features at different frequencies, orientations, and scales with much fewer parameters. For existing loss functions for multi-class image segmentation, there is usually a tradeoff among accuracy, robustness to hyperparameters, and manual weight selections for combining different losses. Therefore, to gain the benefits of using Gabor kernels while keeping the advantage of automatic feature generation in deep learning, we propose a fully trainable Gabor-based convolutional layer where all Gabor parameters are trainable through backpropagation. Furthermore, we propose a loss function based on the Pearson's correlation coefficient, which is accurate, robust to learning rates, and does not require manual weight selections. Experiments on 43 3D brain magnetic resonance images with 19 anatomical structures show that, using the proposed loss function with a proper combination of conventional and Gabor-based kernels, we can train a network with only 1.6 million parameters to achieve an average Dice coefficient of 83%. This size is 44 times smaller than the original V-Net which has 71 million parameters. This paper demonstrates the potentials of using learnable parametric kernels in deep learning for 3D segmentation.

Comments:	This paper was accepted by the International Workshop on Machine Learning in Medical Imaging (MLMI 2022)
Subjects:	Image and Video Processing (eess.IV); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2201.03644 [eess.IV]
	(or arXiv:2201.03644v2 [eess.IV] for this version)
	https://doi.org/10.48550/arXiv.2201.03644
Related DOI:	https://doi.org/10.1007/978-3-031-21014-3_6

Submission history

From: Ken C. L. Wong [view email]
[v1] Mon, 10 Jan 2022 20:55:59 UTC (1,505 KB)
[v2] Thu, 15 Dec 2022 16:24:29 UTC (1,505 KB)

Electrical Engineering and Systems Science > Image and Video Processing

Title:3D Segmentation with Fully Trainable Gabor Kernels and Pearson's Correlation Coefficient

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Electrical Engineering and Systems Science > Image and Video Processing

Title:3D Segmentation with Fully Trainable Gabor Kernels and Pearson's Correlation Coefficient

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators