Semantic segmentation based on multimodal information is used for object detection during autonomous driving at night

Authors

  • Haoqi Wu

DOI:

https://doi.org/10.61173/0s386b39

Keywords:

Autonomous driving, multimodal information semantic segmentation, object detection, Dual U-net network, Dual-way adversarial network

Abstract

Currently, it is difficult for autonomous driving to detect objects accurately enough at night, mainly due to the poor light at night. The laser camera’s captured RGB image information is less effective than during the day, and the image content is seriously affected. The blurred outline of the object and the decreased accuracy of semantic segmentation lead to missed detections. In view of this situation, this study proposes the following solutions: introducing point cloud data collected by lidar during data acquisition, combining it with the image information from the camera, to constitute precise multimodal three-dimensional information. Then, using a dual adversarial network to preprocess the data and U-net semantic segmentation to segment the multimodal 3D information. The simulation experiment is then used to test and evaluate the segmentation effect. After the completion of the training, the object information detection is applied during the autonomous driving of the vehicle. This method, when compared with the general method, has accurate detection, is independent of the intensity of the light, and has the advantage of good universality.

References

[1] Zhao Xiangmo National Key Research and Development Program (2021YFB2501200) team. Research progress in autonomous driving testing and evaluation technology [J]. Transportation engineering,2023,23(06):10-77.DOI:10.19818/ j.cnki.1671-1637.2023.06.002.

[2] Yu Lijiao, Yu Bo, Li Chungeng, et al. Optimize the nighttime human and vehicle detection and identification of convolutional networks and low-resolution thermal imaging [J]. Infrared technology, 2020,42(07):651-659.

[3] Wang Bin, Chen Zhanlong, Wu Liang, et al. Road extraction of high-resolution remote sensing images of U-Net network with both connectivity [J]. Journal of Remote Sensing,2020,24(12):1488-1499.

[4] Jin Yufeng, Tao Ben. A Transformer-based fusion information-enhanced 3D object detection algorithm [J]. Journal of Instrumentation,2023,44(12):297-306.DOI:10.19650/j.cnki. cjsi.J2311940.

[5] Wang Zhongyu, Ni Xianyang, Shang Zhendong. Semantic segmentation of autonomous driving scenes using a convolutional neural network [J]. Optical and precision engineer ing,2019,27(11):2429-2438.

[6] Sun Zhijun, Xue Lei, Xu Yangming, et al. Review of deep learning studies [J]. Computer application research,2012,29(08):2806-2810.

[7] Zhang Xinyu, Gao Hongbo, Zhao Jianhui, et al. Summary of autonomous driving techniques based on deep learning [J]. Journal of Tsinghua University (Natural Science Edition),2018,58(04):438-444.DOI:10.16511/j.cnki. qhdxxb.2018.21.010.

[8] Yang Kang, Chen Li. Research on pedestrian detection based on convolutional neural network in autonomous driving [J]. Computer knowledge and technology,2020,16(25):22-24+30. DOI:10.14004/j.cnki.ckt.2020.2958.

[9] Zhang Shengjian, Mo Zewen. Autonomous driving road detection based on remote sensing images: a histogram equalization strategy [J]. Sensor world,2023,29(08):28-31. DOI:10.16204/j.sw.issn.1006-883X.2023.08.005.

[10] Su Jianmin, Yang Lanxin, Jing Weipeng. Semantic segmentation method for high-resolution remote sensing images based on U-Net [J]. And Computer Engineering and Applicati on,2019,55(07):207-213.

[11] Tang Lili, Liu Gang, Xiao Gang. Infrared and visible light image fusion method based on the two-way cascade confrontation mechanism [J]. Photonics,2021,50(09):321-331.

[12] Zhou Feiyan, Jin Linpeng, Dong Jun. A Review of Convolutional Neural Network Research [J]. The Journal of Computer Science,2017,40(06):1229-1251.

[13] Jing Zhuangwei, Guan Haiyan, Zang Yufu, et al. Review of semantic segmentation research on point clouds based on deep learning [J]. And Computer Science and Exploration,2021,15(01):1-26.

[14] Cao Jingang, Yang Guotian, Yang Xiyong. Deep learning pavement crack detection based on the attention mechanism [J]. The Journal of Computer-Aided Design and Graphics,2020,32(08):1324-1333.

[15] Canina Wang, Chen Yi. Enhanced semantic dual decoder generation model for image repair [J]. Chinese Journal of Image and Graphs,2022,27(10):2994-3009.

Downloads

Published

2024-08-14