Semantic Image Segmentation by Scale-Adaptive Networks

Semantic image segmentation is an important yet unsolved problem. One of the major challenges is the large variability of the object scales. To tackle this scale problem, we propose a Scale-Adaptive Network (SAN) which consists of multiple branches with each one taking charge of the segmentation of...

Ausführliche Beschreibung

Bibliographische Detailangaben
Veröffentlicht in:	IEEE transactions on image processing : a publication of the IEEE Signal Processing Society. - 1992. - 29(2020), 1 vom: 24., Seite 2066-2077
1. Verfasser:	Huang, Zilong (VerfasserIn)
Weitere Verfasser:	Wang, Chunyu, Wang, Xinggang, Liu, Wenyu, Wang, Jingdong
Format:	Online-Aufsatz
Sprache:	English
Veröffentlicht:	2020
Zugriff auf das übergeordnete Werk:	IEEE transactions on image processing : a publication of the IEEE Signal Processing Society
Schlagworte:	Journal Article


LEADER	01000naa a22002652 4500
001	NLM302523561
003	DE-627
005	20231225110913.0
007	cr uuu---uuuuu
008	231225s2020 xx \|\|\|\|\|o 00\| \|\|eng c
024	7		\|a 10.1109/TIP.2019.2941644 \|2 doi
028	5	2	\|a pubmed24n1008.xml
035			\|a (DE-627)NLM302523561
035			\|a (NLM)31647432
040			\|a DE-627 \|b ger \|c DE-627 \|e rakwb
041			\|a eng
100	1		\|a Huang, Zilong \|e verfasserin \|4 aut
245	1	0	\|a Semantic Image Segmentation by Scale-Adaptive Networks
264		1	\|c 2020
336			\|a Text \|b txt \|2 rdacontent
337			\|a ƒaComputermedien \|b c \|2 rdamedia
338			\|a ƒa Online-Ressource \|b cr \|2 rdacarrier
500			\|a Date Completed 07.01.2020
500			\|a Date Revised 07.01.2020
500			\|a published: Print-Electronic
500			\|a Citation Status PubMed-not-MEDLINE
520			\|a Semantic image segmentation is an important yet unsolved problem. One of the major challenges is the large variability of the object scales. To tackle this scale problem, we propose a Scale-Adaptive Network (SAN) which consists of multiple branches with each one taking charge of the segmentation of the objects of a certain range of scales. Given an image, SAN first computes a dense scale map indicating the scale of each pixel which is automatically determined by the size of the enclosing object. Then the features of different branches are fused according to the scale map to generate the final segmentation map. To ensure that each branch indeed learns the features for a certain scale, we propose a scale-induced ground-truth map and enforce a scale-aware segmentation loss for the corresponding branch in addition to the final loss. Extensive experiments over the PASCAL-Person-Part, the PASCAL VOC 2012, and the Look into Person datasets demonstrate that our SAN can handle the large variability of the object scales and outperforms the state-of-the-art semantic segmentation methods
650		4	\|a Journal Article
700	1		\|a Wang, Chunyu \|e verfasserin \|4 aut
700	1		\|a Wang, Xinggang \|e verfasserin \|4 aut
700	1		\|a Liu, Wenyu \|e verfasserin \|4 aut
700	1		\|a Wang, Jingdong \|e verfasserin \|4 aut
773	0	8	\|i Enthalten in \|t IEEE transactions on image processing : a publication of the IEEE Signal Processing Society \|d 1992 \|g 29(2020), 1 vom: 24., Seite 2066-2077 \|w (DE-627)NLM09821456X \|x 1941-0042 \|7 nnns
773	1	8	\|g volume:29 \|g year:2020 \|g number:1 \|g day:24 \|g pages:2066-2077
856	4	0	\|u http://dx.doi.org/10.1109/TIP.2019.2941644 \|3 Volltext
912			\|a GBV_USEFLAG_A
912			\|a SYSFLAG_A
912			\|a GBV_NLM
912			\|a GBV_ILN_350
951			\|a AR
952			\|d 29 \|j 2020 \|e 1 \|b 24 \|h 2066-2077