FCGNet: Foreground and Class Guided Network for human parsing

Cited 0 time in webofscience Cited 0 time in scopus
  • Hit : 88
  • Download : 0
Understanding the inherent hierarchical human structure is key to human parsing. To capture the human- specific characteristic, it is necessary to focus on the spatial and class information corresponding to the foreground (i.e., human) in an image. Inspired by these insights, we introduce two supervision signals, spatial foreground information and existent class information in the image. By utilizing foreground information as guidance, the network is guided to generate a human-focused feature map and capture the pixel-wise hierarchical characteristics by computing correlations between pixels. Furthermore, we guide the network to consider class information in the image at the feature level and capture the class-wise relationship by calculating correlations between channels. Moreover, during the training phase, we prevent the network from misclassifying pixels into confusing classes by providing the existent class information in the image to the network at the prediction level. Our model achieves state-of-the-art performance with significantly reduced parameters and Multiply-Accumulate Operations (MACs) in three public benchmarks.
Publisher
ELSEVIER SCI LTD
Issue Date
2025-01
Language
English
Article Type
Article
Citation

PATTERN RECOGNITION, v.157

ISSN
0031-3203
DOI
10.1016/j.patcog.2024.110879
URI
http://hdl.handle.net/10203/323841
Appears in Collection
EE-Journal Papers(저널논문)
Files in This Item
There are no files associated with this item.

qr_code

  • mendeley

    citeulike


rss_1.0 rss_2.0 atom_1.0