Page 54 - 《软件学报》2026年第5期
P. 54
唐文能 等: 交通场景多模态双阶反馈的三维目标检测方法 1933
[6] Qi CR, Su H, Kaichun M, Guibas LJ. PointNet: Deep learning on point sets for 3D classification and segmentation. In: Proc. of the 2017
IEEE Conf. on Computer Vision and Pattern Recognition. Honolulu: IEEE, 2017. 77–85. [doi: 10.1109/CVPR.2017.16]
[7] Qi CR, Yi L, Su H, Guibas LJ. PointNet++: Deep hierarchical feature learning on point sets in a metric space. In: Proc. of the 31st Int’l
Conf. on Neural Information Processing Systems. Long Beach: Curran Associates Inc., 2017. 5105–5114.
[8] Qi CR, Litany O, He KM, Guibas LJ. Deep Hough voting for 3D object detection in point clouds. In: Proc. of the 2019 IEEE/CVF Int’l
Conf. on Computer Vision. Seoul: IEEE, 2019. 9276–9285. [doi: 10.1109/ICCV.2019.00937]
[9] Shi SS, Wang XG, Li HS. PointRCNN: 3D object proposal generation and detection from point cloud. In: Proc. of the 2019 IEEE/CVF
Conf. on Computer Vision and Pattern Recognition. Long Beach: IEEE, 2019. 770–779. [doi: 10.1109/CVPR.2019.00086]
[10] Sheng HL, Cai SJ, Liu Y, Deng B, Huang JQ, Hua XS, Zhao MJ. Improving 3D object detection with channel-wise Transformer. In:
Proc. of the 2021 IEEE/CVF Int’l Conf. on Computer Vision. Montreal: IEEE, 2021. 2723–2732. [doi: 10.1109/ICCV48922.2021.00274]
[11] Yang ZT, Sun YN, Liu S, Jia JY. 3DSSD: Point-based 3D single stage object detector. In: Proc. of the 2020 IEEE/CVF Conf. on
Computer Vision and Pattern Recognition. Seattle: IEEE, 2020. 11037–11045. [doi: 10.1109/CVPR42600.2020.01105]
[12] Zhang YF, Hu QY, Xu GQ, Ma YX, Wan JW, Guo YL. Not all points are equal: Learning highly efficient point-based detectors for 3D
LiDAR point clouds. In: Proc. of the 2022 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. New Orleans: IEEE, 2022.
18931–18940. [doi: 10.1109/CVPR52688.2022.01838]
[13] Chen C, Chen Z, Zhang J, Tao DC. SASA: Semantics-augmented set abstraction for point-based 3D object detection. In: Proc. of the 36th
AAAI Conf. on Artificial Intelligence. AAAI Press, 2022. 221–229. [doi: 10.1609/aaai.v36i1.19897]
[14] Lang AH, Vora S, Caesar H, Zhou LB, Yang J, Beijbom O. PointPillars: Fast encoders for object detection from point clouds. In: Proc. of
the 2019 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Long Beach: IEEE, 2019. 12689–12697. [doi: 10.1109/CVPR.
2019.01298]
[15] Shi GS, Li RF, Ma C. PillarNet: Real-time and high-performance pillar-based 3D object detection. In: Proc. of the 17th European Conf.
on Computer Vision. Tel Aviv: Springer, 2022. 35–52. [doi: 10.1007/978-3-031-20080-9_3]
[16] Li JY, Luo CX, Yang XD. PillarNeXt: Rethinking network designs for 3D object detection in LiDAR point clouds. In: Proc. of the 2023
IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Vancouver: IEEE, 2023. 17567–17576. [doi: 10.1109/CVPR52729.2023.
01685]
[17] Zhou Y, Tuzel O. VoxelNet: End-to-end learning for point cloud based 3D object detection. In: Proc. of the 2018 IEEE/CVF Conf. on
Computer Vision and Pattern Recognition. Salt Lake City: IEEE, 2018. 4490–4499. [doi: 10.1109/CVPR.2018.00472]
[18] Yan Y, Mao YX, Li B. SECOND: Sparsely embedded convolutional detection. Sensors, 2018, 18(10): 3337. [doi: 10.3390/s18103337]
[19] Duan KW, Bai S, Xie LX, Qi HG, Huang QM, Tian Q. CenterNet: Keypoint triplets for object detection. In: Proc. of the 2019 IEEE/CVF
Int’l Conf. on Computer Vision. Seoul: IEEE, 2019. 6568–6577. [doi: 10.1109/ICCV.2019.00667]
[20] Ge RZ, Ding ZZ, Hu YH, Wang Y, Chen SJ, Huang L, Li Y. AFDet: Anchor free one stage 3D object detection. arXiv:2006.12671, 2020.
[21] He CH, Zeng H, Huang JQ, Hua XS, Zhang L. Structure aware single-stage 3D object detection from point cloud. In: Proc. of the 2020
IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Seattle: IEEE, 2020. 11870–11879. [doi: 10.1109/CVPR42600.2020.
01189]
[22] Zheng W, Tang WL, Chen SJ, Jiang L, Fu CW. CIA-SSD: Confident IoU-aware single-stage object detector from point cloud. In: Proc.
of the 35th AAAI Conf. on Artificial Intelligence. AAAI Press, 2021. 3555–3562. [doi: 10.1609/aaai.v35i4.16470]
[23] Yang HH, Wang WX, Chen MH, Lin BB, He T, Chen H, He XF, Ouyang WL. PVT-SSD: Single-stage 3D object detector with point-
voxel Transformer. In: Proc. of the 2023 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Vancouver: IEEE, 2023.
13476–13487. [doi: 10.1109/CVPR52729.2023.01295]
[24] Shi GS, Li RF, Ma C. Pillar R-CNN for point cloud 3D object detection. arXiv:2302.13301, 2023.
[25] Deng JJ, Shi SS, Li PW, Zhou WG, Zhang YY, Li HQ. Voxel R-CNN: Towards high performance voxel-based 3D object detection. In:
Proc. of the 35th AAAI Conf. on Artificial Intelligence. AAAI Press, 2021. 1201–1209. [doi: 10.1609/aaai.v35i2.16207]
[26] Wu H, Deng JH, Wen CL, Li X, Wang C, Li J. CasA: A cascade attention network for 3-D object detection from LiDAR point clouds.
IEEE Trans. on Geoscience and Remote Sensing, 2022, 60: 5704511. [doi: 10.1109/TGRS.2022.3203163]
[27] Wu H, Wen CL, Li W, Li X, Yang RG, Wang C. Transformation-equivariant 3D object detection for autonomous driving. In: Proc. of the
37th AAAI Conf. on Artificial Intelligence. Washington: AAAI Press, 2023. 2795–2802. [doi: 10.1609/aaai.v37i3.25380]
[28] Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I. Attention is all you need. In: Proc. of the
31st Int’l Conf. on Neural Information Processing Systems. Long Beach: Curran Associates Inc., 2017. 6000–6010.
[29] Wang HY, Shi C, Shi SS, Lei M, Wang S, He D, Schiele B, Wang LW. DSVT: Dynamic sparse voxel Transformer with rotated sets. In:
Proc. of the 2023 IEEE/CVF Conf. on Computer Vision and Pattern Recognition. Vancouver: IEEE, 2023. 13520–13529. [doi: 10.1109/
CVPR52729.2023.01299]

