Page 54 - 《软件学报》2026年第5期
P. 54

唐文能 等: 交通场景多模态双阶反馈的三维目标检测方法                                                     1933


                  [6]   Qi CR, Su H, Kaichun M, Guibas LJ. PointNet: Deep learning on point sets for 3D classification and segmentation. In: Proc. of the 2017
                     IEEE Conf. on Computer Vision and Pattern Recognition. Honolulu: IEEE, 2017. 77–85. [doi: 10.1109/CVPR.2017.16]
                  [7]   Qi CR, Yi L, Su H, Guibas LJ. PointNet++: Deep hierarchical feature learning on point sets in a metric space. In: Proc. of the 31st Int’l
                     Conf. on Neural Information Processing Systems. Long Beach: Curran Associates Inc., 2017. 5105–5114.
                  [8]   Qi CR, Litany O, He KM, Guibas LJ. Deep Hough voting for 3D object detection in point clouds. In: Proc. of the 2019 IEEE/CVF Int’l
                     Conf. on Computer Vision. Seoul: IEEE, 2019. 9276–9285. [doi: 10.1109/ICCV.2019.00937]
                  [9]   Shi SS, Wang XG, Li HS. PointRCNN: 3D object proposal generation and detection from point cloud. In: Proc. of the 2019 IEEE/CVF
                     Conf. on Computer Vision and Pattern Recognition. Long Beach: IEEE, 2019. 770–779. [doi: 10.1109/CVPR.2019.00086]
                 [10]   Sheng HL, Cai SJ, Liu Y, Deng B, Huang JQ, Hua XS, Zhao MJ. Improving 3D object detection with channel-wise Transformer. In:
                     Proc. of the 2021 IEEE/CVF Int’l Conf. on Computer Vision. Montreal: IEEE, 2021. 2723–2732. [doi: 10.1109/ICCV48922.2021.00274]
                 [11]   Yang  ZT,  Sun  YN,  Liu  S,  Jia  JY.  3DSSD:  Point-based  3D  single  stage  object  detector.  In:  Proc.  of  the  2020  IEEE/CVF  Conf.  on
                     Computer Vision and Pattern Recognition. Seattle: IEEE, 2020. 11037–11045. [doi: 10.1109/CVPR42600.2020.01105]
                 [12]   Zhang YF, Hu QY, Xu GQ, Ma YX, Wan JW, Guo YL. Not all points are equal: Learning highly efficient point-based detectors for 3D
                     LiDAR point clouds. In: Proc. of the 2022 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. New Orleans: IEEE, 2022.
                     18931–18940. [doi: 10.1109/CVPR52688.2022.01838]
                 [13]   Chen C, Chen Z, Zhang J, Tao DC. SASA: Semantics-augmented set abstraction for point-based 3D object detection. In: Proc. of the 36th
                     AAAI Conf. on Artificial Intelligence. AAAI Press, 2022. 221–229. [doi: 10.1609/aaai.v36i1.19897]
                 [14]   Lang AH, Vora S, Caesar H, Zhou LB, Yang J, Beijbom O. PointPillars: Fast encoders for object detection from point clouds. In: Proc. of
                     the 2019 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Long Beach: IEEE, 2019. 12689–12697. [doi: 10.1109/CVPR.
                     2019.01298]
                 [15]   Shi GS, Li RF, Ma C. PillarNet: Real-time and high-performance pillar-based 3D object detection. In: Proc. of the 17th European Conf.
                     on Computer Vision. Tel Aviv: Springer, 2022. 35–52. [doi: 10.1007/978-3-031-20080-9_3]
                 [16]   Li JY, Luo CX, Yang XD. PillarNeXt: Rethinking network designs for 3D object detection in LiDAR point clouds. In: Proc. of the 2023
                     IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Vancouver: IEEE, 2023. 17567–17576. [doi: 10.1109/CVPR52729.2023.
                     01685]
                 [17]   Zhou Y, Tuzel O. VoxelNet: End-to-end learning for point cloud based 3D object detection. In: Proc. of the 2018 IEEE/CVF Conf. on
                     Computer Vision and Pattern Recognition. Salt Lake City: IEEE, 2018. 4490–4499. [doi: 10.1109/CVPR.2018.00472]
                 [18]   Yan Y, Mao YX, Li B. SECOND: Sparsely embedded convolutional detection. Sensors, 2018, 18(10): 3337. [doi: 10.3390/s18103337]
                 [19]   Duan KW, Bai S, Xie LX, Qi HG, Huang QM, Tian Q. CenterNet: Keypoint triplets for object detection. In: Proc. of the 2019 IEEE/CVF
                     Int’l Conf. on Computer Vision. Seoul: IEEE, 2019. 6568–6577. [doi: 10.1109/ICCV.2019.00667]
                 [20]   Ge RZ, Ding ZZ, Hu YH, Wang Y, Chen SJ, Huang L, Li Y. AFDet: Anchor free one stage 3D object detection. arXiv:2006.12671, 2020.
                 [21]   He CH, Zeng H, Huang JQ, Hua XS, Zhang L. Structure aware single-stage 3D object detection from point cloud. In: Proc. of the 2020
                     IEEE/CVF  Conf.  on  Computer  Vision  and  Pattern  Recognition.  Seattle:  IEEE,  2020.  11870–11879.  [doi:  10.1109/CVPR42600.2020.
                     01189]
                 [22]   Zheng W, Tang WL, Chen SJ, Jiang L, Fu CW. CIA-SSD: Confident IoU-aware single-stage object detector from point cloud. In: Proc.
                     of the 35th AAAI Conf. on Artificial Intelligence. AAAI Press, 2021. 3555–3562. [doi: 10.1609/aaai.v35i4.16470]
                 [23]   Yang HH, Wang WX, Chen MH, Lin BB, He T, Chen H, He XF, Ouyang WL. PVT-SSD: Single-stage 3D object detector with point-
                     voxel  Transformer.  In:  Proc.  of  the  2023  IEEE/CVF  Conf.  on  Computer  Vision  and  Pattern  Recognition.  Vancouver:  IEEE,  2023.
                     13476–13487. [doi: 10.1109/CVPR52729.2023.01295]
                 [24]   Shi GS, Li RF, Ma C. Pillar R-CNN for point cloud 3D object detection. arXiv:2302.13301, 2023.
                 [25]   Deng JJ, Shi SS, Li PW, Zhou WG, Zhang YY, Li HQ. Voxel R-CNN: Towards high performance voxel-based 3D object detection. In:
                     Proc. of the 35th AAAI Conf. on Artificial Intelligence. AAAI Press, 2021. 1201–1209. [doi: 10.1609/aaai.v35i2.16207]
                 [26]   Wu H, Deng JH, Wen CL, Li X, Wang C, Li J. CasA: A cascade attention network for 3-D object detection from LiDAR point clouds.
                     IEEE Trans. on Geoscience and Remote Sensing, 2022, 60: 5704511. [doi: 10.1109/TGRS.2022.3203163]
                 [27]   Wu H, Wen CL, Li W, Li X, Yang RG, Wang C. Transformation-equivariant 3D object detection for autonomous driving. In: Proc. of the
                     37th AAAI Conf. on Artificial Intelligence. Washington: AAAI Press, 2023. 2795–2802. [doi: 10.1609/aaai.v37i3.25380]
                 [28]   Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I. Attention is all you need. In: Proc. of the
                     31st Int’l Conf. on Neural Information Processing Systems. Long Beach: Curran Associates Inc., 2017. 6000–6010.
                 [29]   Wang HY, Shi C, Shi SS, Lei M, Wang S, He D, Schiele B, Wang LW. DSVT: Dynamic sparse voxel Transformer with rotated sets. In:
                     Proc. of the 2023 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Vancouver: IEEE, 2023. 13520–13529. [doi: 10.1109/
                     CVPR52729.2023.01299]
   49   50   51   52   53   54   55   56   57   58   59