Pyramid Scene Parsing Network

2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)(2017)

引用 13839|浏览672
暂无评分
摘要
Scene parsing is challenging for unrestricted open vocabulary and diverse scenes. In this paper, we exploit the capability of global context information by different-region-based context aggregation through our pyramid pooling module together with the proposed pyramid scene parsing network (PSPNet). Our global prior representation is effective to produce good quality results on the scene parsing task, while PSPNet provides a superior framework for pixel-level prediction tasks. The proposed approach achieves state-of-the-art performance on various datasets. It came first in ImageNet scene parsing challenge 2016, PASCAL VOC 2012 benchmark and Cityscapes benchmark. A single PSPNet yields new record of mIoU accuracy 85.4% on PASCAL VOC 2012 and accuracy 80.2% on Cityscapes.
更多
查看译文
关键词
pyramid scene parsing network,ImageNet scene parsing challenge,scene parsing task,global prior representation,PSPNet,pyramid pooling module,global context information,unrestricted open vocabulary
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要