Cargando…

Instance segmentation convolutional neural network based on multi-scale attention mechanism

Instance segmentation is more challenging and difficult than object detection and semantic segmentation. It paves the way for the realization of a complete scene understanding, and has been widely used in robotics, automatic driving, medical care, and other aspects. However, there are some problems...

Descripción completa

Detalles Bibliográficos
Autores principales: Gaihua, Wang, Jinheng, Lin, Lei, Cheng, Yingying, Dai, Tianlun, Zhang
Formato: Online Artículo Texto
Lenguaje:English
Publicado: Public Library of Science 2022
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8794127/
https://www.ncbi.nlm.nih.gov/pubmed/35085359
http://dx.doi.org/10.1371/journal.pone.0263134
_version_ 1784640760704925696
author Gaihua, Wang
Jinheng, Lin
Lei, Cheng
Yingying, Dai
Tianlun, Zhang
author_facet Gaihua, Wang
Jinheng, Lin
Lei, Cheng
Yingying, Dai
Tianlun, Zhang
author_sort Gaihua, Wang
collection PubMed
description Instance segmentation is more challenging and difficult than object detection and semantic segmentation. It paves the way for the realization of a complete scene understanding, and has been widely used in robotics, automatic driving, medical care, and other aspects. However, there are some problems in instance segmentation methods, such as the low detection efficiency for low-resolution objects and the slow detection speed of images with complex backgrounds. To solve these problems, this paper proposes an instance segmentation method with multi-scale attention, which is called a Hybrid Kernel Mask R-CNN. Firstly, the hybrid convolution kernel is constructed by combining different kernels and groups, which can complement each other to extract rich information. Secondly, a multi-scale attention mechanism is designed by assign weights to different convolution kernels, which can retain more important information. After the introduction of our strategy, the network is more inclined to focus on the low-resolution objects in the image. The proposed method achieves the best accuracy over the anchor-based method. To verify the universality of the model, we test Hybrid Kernel Mask R-CNN on Balloon, xBD and COCO datasets. The test results exceed the state of art methods. And the visualization results show our method can extract low-resolution objects effectively.
format Online
Article
Text
id pubmed-8794127
institution National Center for Biotechnology Information
language English
publishDate 2022
publisher Public Library of Science
record_format MEDLINE/PubMed
spelling pubmed-87941272022-01-28 Instance segmentation convolutional neural network based on multi-scale attention mechanism Gaihua, Wang Jinheng, Lin Lei, Cheng Yingying, Dai Tianlun, Zhang PLoS One Research Article Instance segmentation is more challenging and difficult than object detection and semantic segmentation. It paves the way for the realization of a complete scene understanding, and has been widely used in robotics, automatic driving, medical care, and other aspects. However, there are some problems in instance segmentation methods, such as the low detection efficiency for low-resolution objects and the slow detection speed of images with complex backgrounds. To solve these problems, this paper proposes an instance segmentation method with multi-scale attention, which is called a Hybrid Kernel Mask R-CNN. Firstly, the hybrid convolution kernel is constructed by combining different kernels and groups, which can complement each other to extract rich information. Secondly, a multi-scale attention mechanism is designed by assign weights to different convolution kernels, which can retain more important information. After the introduction of our strategy, the network is more inclined to focus on the low-resolution objects in the image. The proposed method achieves the best accuracy over the anchor-based method. To verify the universality of the model, we test Hybrid Kernel Mask R-CNN on Balloon, xBD and COCO datasets. The test results exceed the state of art methods. And the visualization results show our method can extract low-resolution objects effectively. Public Library of Science 2022-01-27 /pmc/articles/PMC8794127/ /pubmed/35085359 http://dx.doi.org/10.1371/journal.pone.0263134 Text en © 2022 Gaihua et al https://creativecommons.org/licenses/by/4.0/This is an open access article distributed under the terms of the Creative Commons Attribution License (https://creativecommons.org/licenses/by/4.0/) , which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited.
spellingShingle Research Article
Gaihua, Wang
Jinheng, Lin
Lei, Cheng
Yingying, Dai
Tianlun, Zhang
Instance segmentation convolutional neural network based on multi-scale attention mechanism
title Instance segmentation convolutional neural network based on multi-scale attention mechanism
title_full Instance segmentation convolutional neural network based on multi-scale attention mechanism
title_fullStr Instance segmentation convolutional neural network based on multi-scale attention mechanism
title_full_unstemmed Instance segmentation convolutional neural network based on multi-scale attention mechanism
title_short Instance segmentation convolutional neural network based on multi-scale attention mechanism
title_sort instance segmentation convolutional neural network based on multi-scale attention mechanism
topic Research Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC8794127/
https://www.ncbi.nlm.nih.gov/pubmed/35085359
http://dx.doi.org/10.1371/journal.pone.0263134
work_keys_str_mv AT gaihuawang instancesegmentationconvolutionalneuralnetworkbasedonmultiscaleattentionmechanism
AT jinhenglin instancesegmentationconvolutionalneuralnetworkbasedonmultiscaleattentionmechanism
AT leicheng instancesegmentationconvolutionalneuralnetworkbasedonmultiscaleattentionmechanism
AT yingyingdai instancesegmentationconvolutionalneuralnetworkbasedonmultiscaleattentionmechanism
AT tianlunzhang instancesegmentationconvolutionalneuralnetworkbasedonmultiscaleattentionmechanism