基于图卷积网络的少镜头分割原型混合模型

Zhibo Gu, Zhiming Luo, Min Huang, Yuanzheng Cai, Shaozi Li
{"title":"基于图卷积网络的少镜头分割原型混合模型","authors":"Zhibo Gu, Zhiming Luo, Min Huang, Yuanzheng Cai, Shaozi Li","doi":"10.1109/ITME53901.2021.00028","DOIUrl":null,"url":null,"abstract":"Over the past few years, deep convolutional neural networks (CNNs) based semantic segmentation methods reached the state-of-the-art performance. To train a model with the ability to know a concept, a lot of pixel level annotated images are required, which is time consuming and hard to cover unseen object categories. Thus, few-shot semantic segmentation has been developed to implement segmentation with a few annotation images. In this paper, we proposed a novel prototype mixing model for few shot segmentation. Different with other works which only produce prototypes form support set, our proposed model learn a group of concept-specific prototypes from support set and then generate prototypes from query set. With prototypes from both query set and support set, we proposed a GCN(Graphic Convolutional Network) module to generate mixing prototypes for better utilizing of informations from different categories. We also proposed a clustering module to produce multi-prototypes for representing different parts of a single semantic class, which reach better performance than single prototype. Our model achieve 48.8% and 55.9%mIoU score on PASCAL-5i for 1-shot and 5-shot settings respectively.","PeriodicalId":6774,"journal":{"name":"2021 11th International Conference on Information Technology in Medicine and Education (ITME)","volume":"37 1","pages":"86-90"},"PeriodicalIF":0.0000,"publicationDate":"2021-11-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"A Graph-Convolutional-Network based Prototype Mixing Model for Few-shot Segmentation\",\"authors\":\"Zhibo Gu, Zhiming Luo, Min Huang, Yuanzheng Cai, Shaozi Li\",\"doi\":\"10.1109/ITME53901.2021.00028\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Over the past few years, deep convolutional neural networks (CNNs) based semantic segmentation methods reached the state-of-the-art performance. To train a model with the ability to know a concept, a lot of pixel level annotated images are required, which is time consuming and hard to cover unseen object categories. Thus, few-shot semantic segmentation has been developed to implement segmentation with a few annotation images. In this paper, we proposed a novel prototype mixing model for few shot segmentation. Different with other works which only produce prototypes form support set, our proposed model learn a group of concept-specific prototypes from support set and then generate prototypes from query set. With prototypes from both query set and support set, we proposed a GCN(Graphic Convolutional Network) module to generate mixing prototypes for better utilizing of informations from different categories. We also proposed a clustering module to produce multi-prototypes for representing different parts of a single semantic class, which reach better performance than single prototype. Our model achieve 48.8% and 55.9%mIoU score on PASCAL-5i for 1-shot and 5-shot settings respectively.\",\"PeriodicalId\":6774,\"journal\":{\"name\":\"2021 11th International Conference on Information Technology in Medicine and Education (ITME)\",\"volume\":\"37 1\",\"pages\":\"86-90\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2021-11-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2021 11th International Conference on Information Technology in Medicine and Education (ITME)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/ITME53901.2021.00028\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2021 11th International Conference on Information Technology in Medicine and Education (ITME)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/ITME53901.2021.00028","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 0

摘要

在过去的几年里,基于深度卷积神经网络(cnn)的语义分割方法达到了最先进的性能。为了训练一个具有认识概念能力的模型,需要大量的像素级注释图像,这是耗时的,而且很难覆盖看不见的对象类别。因此,利用少量标注图像实现语义分割的方法被开发出来。本文提出了一种用于小镜头分割的原型混合模型。与以往只从支持集中生成原型的方法不同,本文提出的模型从支持集中学习一组特定概念的原型,然后从查询集中生成原型。基于查询集和支持集的原型,我们提出了一个图形卷积网络(GCN)模块来生成混合原型,以便更好地利用来自不同类别的信息。我们还提出了一个聚类模块来生成多个原型来表示单个语义类的不同部分,从而达到比单个原型更好的性能。我们的模型在1次射击和5次射击设置下分别在PASCAL-5i上获得48.8%和55.9%的miou得分。
本文章由计算机程序翻译,如有差异,请以英文原文为准。
A Graph-Convolutional-Network based Prototype Mixing Model for Few-shot Segmentation
Over the past few years, deep convolutional neural networks (CNNs) based semantic segmentation methods reached the state-of-the-art performance. To train a model with the ability to know a concept, a lot of pixel level annotated images are required, which is time consuming and hard to cover unseen object categories. Thus, few-shot semantic segmentation has been developed to implement segmentation with a few annotation images. In this paper, we proposed a novel prototype mixing model for few shot segmentation. Different with other works which only produce prototypes form support set, our proposed model learn a group of concept-specific prototypes from support set and then generate prototypes from query set. With prototypes from both query set and support set, we proposed a GCN(Graphic Convolutional Network) module to generate mixing prototypes for better utilizing of informations from different categories. We also proposed a clustering module to produce multi-prototypes for representing different parts of a single semantic class, which reach better performance than single prototype. Our model achieve 48.8% and 55.9%mIoU score on PASCAL-5i for 1-shot and 5-shot settings respectively.
求助全文
通过发布文献求助,成功后即可免费获取论文全文。 去求助
来源期刊
自引率
0.00%
发文量
0
×
引用
GB/T 7714-2015
复制
MLA
复制
APA
复制
导出至
BibTeX EndNote RefMan NoteFirst NoteExpress
×
提示
您的信息不完整,为了账户安全,请先补充。
现在去补充
×
提示
您因"违规操作"
具体请查看互助需知
我知道了
×
提示
确定
请完成安全验证×
copy
已复制链接
快去分享给好友吧!
我知道了
右上角分享
点击右上角分享
0
联系我们:info@booksci.cn Book学术提供免费学术资源搜索服务,方便国内外学者检索中英文文献。致力于提供最便捷和优质的服务体验。 Copyright © 2023 布克学术 All rights reserved.
京ICP备2023020795号-1
ghs 京公网安备 11010802042870号
Book学术文献互助
Book学术文献互助群
群 号:481959085
Book学术官方微信