[1] 方向,马楠,韩义恒,等. 多模态信息质量评价与技术挑战 [J]. 中国科学:技术科学, 2026, 56 (6): 1626-1632.
Fang X, Ma N, Han Y H, et al. Quality evaluation and technical challenges of multimodal information[J]. Science China: Technological Sciences, 2026: 1-7.
[2] Ma N, Pan J, Liu Y, et al. Embodied Interactive Intelligence Towards Autonomous Driving[J]. Engineering, 2026, 59: 337-351.
[3] Li X, Han Y, Ma N.Autonomous Tomato Harvesting With Top–Down Fusion Network for Limited Data[J].IEEE Transactions on Robotics,2025,41:3609- 3628.
[4] Ma N,Wu Z X,Feng Y F, et al. Multi-View Time-Series Hypergraph Neural Network for Action Recognition [J].IEEE Transactions on Image Processing, 2024, 33:3301-3313.
[5] Ma N, Wu Z, Li K, et al. Interactive Cognition of Self-driving: A Multidimensional Analysis Model and Implementation[J]. Research, 2025, 8.
[6] 赵江平,王欣然,吴立舟.改进YOLOv5s的路面坑槽目标检测模型[J].中国安全科学学报,2025,35(01):67-74.
Zhao J P, Wang X R, Wu L Z. Improved YOLOv5s-based road pothole detection model[J]. China Safety Science Journal, 2025, 35(1): 67-74.
[7] Mousavian A, Angueira J,Jafari D, et al. BEVDetNet: Bird’s-eye view LiDAR point cloud based real-time 3D object detection[J].IEEE Robotics and Automation Letters,2021,6(2):2947-2954.
[8] Li D,Furukawa T.Global vision-based reconstruction of three-dimensional road surfaces using adaptive extended Kalman filter[C]//Proceedings of the IEEE International Conference on Robotics and Automation(ICRA).Montreal, Canada:IEEE,2019:3860 -3866.
[9] Min C, Xiao L, Zhao D, et al. UniScene: Multi-Camera Unified Pre-training via 3D Scene Reconstruction for Autonomous Driving[EB/OL]. (2023-05-30). https://arxiv. org/abs/2305.18829.
[10] Li H,Wang J,Zhang Y,et al.Delving into the devils of bird’s-eye-view perception:A review,evaluation,and recipe[J].IEEE Transactions on Pattern Analysis and Machine Intelligence,2024,46(4):2151 -2170.
[11] Philion J,Fahim A.Lift,splat,shoot:Encoding images from arbitrary camera rigs by implicitly unprojecting to 3D[C]//Proceedings of the European Conference on Computer Vision(ECCV).Glasgow,UK:Springer,2020:194 -210.
[12] Wang J,Liu H,Chen Q,et al. Toward robust LiDAR-camera fusion in BEV space via mutual deformable attention and temporal aggregation[J].IEEE Transactions on Circuits and Systems for Video Technology,2024,34(7):5753 -5764.
[13] Wang Y ,Mao Q ,Zhu X,et al. Dynamic fusion of LiDAR and camera for 3D object detection in autonomous driving[J].IEEE Transactions on Multimedia,2022,24: 4567–4578.
[14] Song Z,Liu Y,Chen J,et al.GraphBEV:Towards robust BEV feature alignment for multi-modal 3D object detection[EB/OL]. (2024-03-18). https:/arxiv org/abs/ 2403.11848.
[15] Radford A,Kim J W,Hallacy C,et al.Learning transferable visual models from natural language supervision[C]// Proceedings of the International Conference on Machine Learning(ICML).Virtual:PMLR,2021:8748 -8763.
[16] Wang L, Chen Y, Zhang Z, et al. Multi-modal 3D object detection in autonomous driving: A survey and taxonomy[J]. IEEE Transactions on Intelligent Vehicles, 2023, 8(7): 3781-3798.
[17] 李想,王青正.基于车载激光点云数据的车辆行驶路面坑槽检测方法[J].激光杂志,2025,46(3): 227-232.
Li X, Wang Q. A method for detecting road surface potholes based on vehicle-mounted laser point cloud data[J].Journal of Lasers,2025,46(3): 227-232.
[18] Ye W,Jinag W,Tong Z,et al.Convolutional neural network for pothole detection in asphalt pavement[J].Road Materials and Pavement Design,2021,22(1):42-58.
[19] Park S S,Tran V T,Lee D E.Application of various YOLO models for computer-vision-based real-time pothole detection[J].Applied Sciences,2021,11(20):11229.
[20] Guo S,Wang H,Li Y,et al.UDTIRI:An online open-source intelligent road inspection benchmark suite[J].IEEE Transactions on Intelligent Transportation Systems,2024, 25(8):9920 -9931.
[21] 郭奇,李明鸿,赵于前,等.双向注意力特征调制的路面坑槽分割方法[J/OL].计算机工程与应用,2026,58(1):1-10.
Guo Q,Li M,Zhao Y,et al.Road pothole segmentation method based on bidirectional attention feature modulation[J/OL].Computer Engineering and Applications,2026,58(1):1-10.
[22] 金政北,金贝贝,宋晓辉,等.改进YOLOv10算法及其在路面坑洼检测中的应用[J/OL].计算机应用与软件,2026,43(1):1-9.
Jin Z,Jin B,Song X,et al.Improved YOLOv10 algorithm and its application in road pothole detection [J/OL].Computer Applications and Software,2026,43(1): 1-9.
[23] Li J,Zhang Y,Yun P,et al.RoadFormer:Duplex Transformer for RGB-normal semantic road scene parsing[J].IEEE Transactions on Intelligent Vehicles,2024,9(7):5163 -5172.
[24] Zhao T, Xu C, Ding M, et al. RSRD:A Road Surface Reconstruction Dataset and Benchmark for Safe and Comfortable Autonomous Driving[EB/OL].(2023-10-03). https://arxiv.org/abs/2310.02262.
[25] Khan A A,Shao J,Rao Y,et al.LRDNet:Lightweight LiDAR aided cascaded feature pools for free road space detection[J].IEEE Transactions on Multimedia,2025,27: 652-664.
[26] Fan R,Wang H,Cai P,et al.SNE-RoadSeg:Incorporating surface normal information into semantic segmentation for accurate freespace detection [C]//Proceedings of the European Conference on Computer Vision (ECCV). Glasgow, UK: Springer, 2020: 340-356.
[27] Wang H,Fan R,Cai P,et al.SNE-RoadSeg+:Rethinking depth-normal translation and deep supervision for freespace detection[C]//2021 IEEE/RSJ International Conference on Intelligent Robots and Systems(IROS). Prague,Czech Republic:IEEE,2021:1140–1145.
[28] Huang J,Li J,Jia N,et al. RoadFormer+:Delivering RGB-X scene parsing through scale-aware information decoupling and advanced heterogeneous feature fusion[J].IEEE Transactions on Intelligent Vehicles,2024,10: 3156-3165.
[29] Sun Y, Zuo W,Liu M,et al.RTFNet:RGB-thermal fusion network for semantic segmentation of urban scenes[J].IEEE Robotics and Automation Letters,2019, 4(3):2576-2583
[30] Chang J R, Chen Y S.Pyramid stereo matching network[C] //Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Salt Lake City, UT, USA: IEEE, 2018: 5410-5418.
[31] Shen Z, Dai Y, Rao Z. CFNet: Cascade and Fused Cost Volume for Robust Stereo Matching[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). Virtual: IEEE, 2021: 13901-13910.
[32] Xu G, Cheng J, Guo P, et al. Attention Concatenation Volume for Accurate and Efficient Stereo Matching[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). New Orleans, LA, USA: IEEE, 2022: 12971-12980.
[33] Zhao T, Ding M, Zhan W, et al.Depth-aware Volume Attention for Texture-less Stereo Matching[EB/OL]. (2024-02-13). https://arxiv.org/abs/2402.08931.
[34] Zhao T, Yang L, Xie Y, et al. RoadBEV: Road Surface Reconstruction in Bird’s Eye View[J].IEEE Transactions on Intelligent Transportation Systems,2024,25(11):19088- 19099.
[35] Sun L, Zang H, Yin W. Pseudo-LiDAR Based Road Detection[J]. IEEE Transactions on Circuits and Systems for Video Technology, 2022, 32(8): 5386-5398.
[36] Wang H, Fan R, Sun Y, et al. Applying surface normal information in drivable area and road anomaly detection for ground mobile robots[C]//2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). IEEE, 2020: 2706-2711.
[37] Wang H, Fan R, Sun Y, et al. Dynamic fusion module evolves drivable area and road anomaly detection: A benchmark and algorithms[J]. IEEE transactions on cybernetics, 2021, 52(10): 10750-10760.
|