File Download

There are no files associated with this item.

  Links for fulltext
     (May Require Subscription)
Supplementary

Conference Paper: Polygon-free: Unconstrained Scene Text Detection with Box Annotations

TitlePolygon-free: Unconstrained Scene Text Detection with Box Annotations
Authors
Issue Date2022
PublisherIEEE.
Citation
29th IEEE International Conference on Image Processing (ICIP), Bordeaux, France, 16-19 October, 2022. In Proceedings of IEEE International Conference on Image Processing (ICIP) 2022 How to Cite?
AbstractAlthough a polygon is a more accurate representation than an upright bounding box for text detection, the annotations of polygons are extremely expensive and challenging. Unlike existing works that employ fully-supervised training with polygon annotations, this study proposes an unconstrained text detection system termed Polygon-free (PF), in which most existing polygon-based text detectors ( e.g., PSENet [33],DB [16]) are trained with only upright bounding box annotations. Our core idea is to transfer knowledge from synthetic data to real data to enhance the supervision information of upright bounding boxes. This is made pos-sible with a simple segmentation network, namely Skeleton Attention Segmentation Network (SASN), that includes three vital components ( i.e., channel attention, spatial attention and skeleton attention map) and one soft cross-entropy loss. Experiments demonstrate that the proposed Polygon-free system can combine general detectors ( e.g., EAST, PSENet, DB) to yield surprisingly high-quality pixel-level results with only upright bounding box annotations on a variety of datasets ( e.g., ICDAR2019-Art, TotalText, IC-DAR2015). For example, without using polygon annotations, PSENet achieves an 80.5% F-score on TotalText [3] (vs. 80.9% of fully supervised counterpart), 31.1% better than training directly with upright bounding box annotations, and saves 80%+ labeling costs. We hope that PF can provide a new perspective for text detection to reduce the labeling costs.
Persistent Identifierhttp://hdl.handle.net/10722/315807

 

DC FieldValueLanguage
dc.contributor.authorWu, W-
dc.contributor.authorXie, E-
dc.contributor.authorZhang, R-
dc.contributor.authorWang, W-
dc.contributor.authorLuo, P-
dc.contributor.authorZhou, H-
dc.date.accessioned2022-08-19T09:04:48Z-
dc.date.available2022-08-19T09:04:48Z-
dc.date.issued2022-
dc.identifier.citation29th IEEE International Conference on Image Processing (ICIP), Bordeaux, France, 16-19 October, 2022. In Proceedings of IEEE International Conference on Image Processing (ICIP) 2022-
dc.identifier.urihttp://hdl.handle.net/10722/315807-
dc.description.abstractAlthough a polygon is a more accurate representation than an upright bounding box for text detection, the annotations of polygons are extremely expensive and challenging. Unlike existing works that employ fully-supervised training with polygon annotations, this study proposes an unconstrained text detection system termed Polygon-free (PF), in which most existing polygon-based text detectors ( e.g., PSENet [33],DB [16]) are trained with only upright bounding box annotations. Our core idea is to transfer knowledge from synthetic data to real data to enhance the supervision information of upright bounding boxes. This is made pos-sible with a simple segmentation network, namely Skeleton Attention Segmentation Network (SASN), that includes three vital components ( i.e., channel attention, spatial attention and skeleton attention map) and one soft cross-entropy loss. Experiments demonstrate that the proposed Polygon-free system can combine general detectors ( e.g., EAST, PSENet, DB) to yield surprisingly high-quality pixel-level results with only upright bounding box annotations on a variety of datasets ( e.g., ICDAR2019-Art, TotalText, IC-DAR2015). For example, without using polygon annotations, PSENet achieves an 80.5% F-score on TotalText [3] (vs. 80.9% of fully supervised counterpart), 31.1% better than training directly with upright bounding box annotations, and saves 80%+ labeling costs. We hope that PF can provide a new perspective for text detection to reduce the labeling costs.-
dc.languageeng-
dc.publisherIEEE.-
dc.relation.ispartofProceedings of IEEE International Conference on Image Processing (ICIP) 2022-
dc.rightsProceedings of IEEE International Conference on Image Processing (ICIP). Copyright © IEEE.-
dc.titlePolygon-free: Unconstrained Scene Text Detection with Box Annotations-
dc.typeConference_Paper-
dc.identifier.emailLuo, P: pluo@hku.hk-
dc.identifier.authorityLuo, P=rp02575-
dc.identifier.doi10.48550/arXiv.2011.13307-
dc.identifier.hkuros335610-
dc.publisher.placeUnited States-

Export via OAI-PMH Interface in XML Formats


OR


Export to Other Non-XML Formats