Page 488 - 《软件学报》2026年第2期
P. 488

童彪 等: 基于边缘增强的宽解码器显著性目标检测方法                                                       967


                  [5]   Simard PY, Steinkraus D, Platt JC. Best practices for convolutional neural networks applied to visual document analysis. In: Proc. of the
                     7th Int’l Conf. on Document Analysis and Recognition. Edinburgh: IEEE, 2003. 958–963. [doi: 10.1109/ICDAR.2003.1227801]
                  [6]   Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I. Attention is all you need. In: Proc. of the
                     31st  Int’l  Conf.  on  Neural  Information  Processing  Systems.  Long  Beach:  Curran  Associates  Inc.,  2017.  6000–6010.  [doi:  10.5555/
                     3295222.3295349]
                  [7]   Dosovitskiy  A,  Beyer  L,  Kolesnikov  A,  Weissenborn  D,  Zhai  XH,  Unterthiner  T,  Dehghani  M,  Minderer  M,  Heigold  G,  Gelly  S,
                     Uszkoreit J, Houlsby N. An image is worth 16×16 words: Transformers for image recognition at scale. In: Proc. of the 9th Int’l Conf. on
                     Learning Representations. 2021.
                  [8]   Kyurkchiev N, Svetoslav M. Sigmoid Functions: Some Approximation and Modelling Aspects. Saarbrucken: LAP LAMBERT Academic
                     Publishing, 2015. [doi: 10.11145/j.bmc.2015.03.081]
                  [9]   Chen LY, Li SB, Bai Q, Yang J, Jiang SL, Miao YM. Review of image classification algorithms based on convolutional neural networks.
                     Remote Sensing, 2021, 13(22): 4712. [doi: 10.3390/rs13224712]
                 [10]   Zou ZX, Chen KY, Shi ZW, Guo YH, Ye JP. Object detection in 20 years: A survey. Proc. of the IEEE, 2023, 111(3): 257–276. [doi: 10.
                     1109/JPROC.2023.3238524]
                 [11]   Dharampal, Mutneja V. Methods of image edge detection: A review. Journal of Electrical & Electronic Systems, 2015, 4(2): 1000150.
                     [doi: 10.4172/2332-0796.1000150]
                 [12]   Prewitt JMS. Object enhancement and extraction. Picture Processing and Psychopictorics, 1970, 10: 15–19.
                 [13]   Yu ZT, Zhao CX, Wang ZZ, Qin YX, Su Z, Li XB, Zhou F, Zhao GY. Searching central difference convolutional networks for face anti-
                     spoofing. In: Proc. of the 2020 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Seattle: IEEE, 2020. 5294–5304. [doi: 10.
                     1109/CVPR42600.2020.00534]
                 [14]   Su Z, Liu WZ, Yu ZT, Hu DW, Liao Q, Tian Q, Pietikäinen M, Liu L. Pixel difference networks for efficient edge detection. In: Proc. of
                     the 2021 IEEE/CVF Int’l Conf. on Computer Vision. Montreal: IEEE, 2021. 5097–5107. [doi: 10.1109/ICCV48922.2021.00507]
                 [15]   Ronneberger O, Fischer P, Brox T. U-net: Convolutional networks for biomedical image segmentation. In: Proc. of the 18th Int’l Conf. on
                     Medical Image Computing and Computer-assisted Intervention (MICCAI 2015). Munich: Springer, 2015. 234–241. [doi: 10.1007/978-3-
                     319-24574-4_28]
                 [16]   Tian  Z,  He  T,  Shen  CH,  Yan  YL.  Decoders  matter  for  semantic  segmentation:  Data-dependent  decoding  enables  flexible  feature
                     aggregation. In: Proc. of the 2019 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Long Beach: IEEE, 2019. 3121–3130.
                     [doi: 10.1109/CVPR.2019.00324]
                 [17]   Ding XH, Zhang XY, Han JG, Ding GG. Scaling up your kernels to 31×31: Revisiting large kernel design in CNNs. In: Proc. of the 2022
                     IEEE/CVF Conf. on Computer Vision and Pattern Recognition. New Orleans: IEEE, 2022. 11953–11965. [doi: 10.1109/CVPR52688.
                     2022.01166]
                 [18]   Yu F, Koltun V. Multi-scale context aggregation by dilated convolutions. In: Proc. of the 4th Int’l Conf. on Learning Representations.
                     2016.
                 [19]   Ioffe S, Szegedy C. Batch normalization: Accelerating deep network training by reducing internal covariate shift. In: Proc. of the 32nd Int’l
                     Conf. on Machine Learning, Vol. 37. 2015. 448–456. [doi: 10.5555/3045118.3045167]
                 [20]   Wang LJ, Lu HC, Wang YF, Feng MY, Wang D, Yin BC, Ruan X. Learning to detect salient objects with image-level supervision. In:
                     Proc. of the 2017 IEEE Conf. on Computer Vision and Pattern Recognition. Honolulu: IEEE, 2017. 3796–3805. [doi: 10.1109/CVPR.
                     2017.404]
                 [21]   Yang C, Zhang LH, Lu HC, Ruan X, Yang MH. Saliency detection via graph-based manifold ranking. In: Proc. of the 2013 IEEE Conf.
                     on Computer Vision and Pattern Recognition. Portland: IEEE, 2013. 3166–3173. [doi: 10.1109/CVPR.2013.407]
                 [22]   Zeng Y, Zhang PP, Lin Z, Zhang JM, Lu HC. Towards high-resolution salient object detection. In: Proc. of the 2019 IEEE/CVF Int’l
                     Conf. on Computer Vision. Seoul: IEEE, 2019. 7233–7242. [doi: 10.1109/ICCV.2019.00733]
                 [23]   Perazzi F, Wang O, Gross M, Sorkine-Hornung A. Fully connected object proposals for video segmentation. In: Proc. of the 2015 IEEE
                     Int’l Conf. on Computer Vision. Santiago: IEEE, 2015. 3227–3234. [doi: 10.1109/ICCV.2015.369]
                 [24]   Sokolova  M,  Japkowicz  N,  Szpakowicz  S.  Beyond  accuracy,  F-score  and  ROC:  A  family  of  discriminant  measures  for  performance
                     evaluation.  In:  Proc.  of  the  19th  Australian  Joint  Conf.  on  Artificial  Intelligence.  Hobart:  Springer,  2006.  1015–1021.  [doi:  10.1007/
                     11941439_114]
                 [25]   Fan DP, Cheng MM, Liu Y, Li T, Borji A. Structure-measure: A new way to evaluate foreground maps. In: Proc. of the 2017 IEEE Int’l
                     Conf. on Computer Vision. Venice: IEEE, 2017. 4558–4567. [doi: 10.1109/ICCV.2017.487]
                 [26]   Fan DP, Gong C, Cao Y, Ren B, Cheng MM, Borji A. Enhanced-alignment measure for binary foreground map evaluation. In: Proc. of
   483   484   485   486   487   488   489   490