Electronics and Electrical Engineering and Control

Robustness design and analysis of airborne visual perception based on deep ensemble learning

  • Zan MA ,
  • Tongjie ZHANG ,
  • Jie BAI ,
  • Yong CHEN ,
  • Yi TIAN
Expand
  • 1.College of Safety Science and Engineering,Civil Aviation University of China,Tianjin 300300,China
    2.Key Laboratory of Civil Aircraft Airworthiness Certification Technology,Civil Aviation University of China,Tianjin 300300,China
    3.Shanghai Aircraft Design and Research Institute,Commercial Aircraft Corporation of China,Shanghai 200216,China
E-mail: jbai@cauc.edu.cn

Received date: 2025-10-13

  Revised date: 2025-11-10

  Accepted date: 2025-12-02

  Online published: 2025-12-25

Supported by

National Key Research and Development Program of China(2022YFB3904300)

Abstract

The visual perception function based on machine learning is crucial for enhancing situational awareness or autonomous flight capabilities of aircraft in complex environments, and its performance has significant impact on flight safety. However, the inherent probabilistic nature of machine learning techniques poses substantial challenges to meeting airworthiness safety objectives, thereby hindering their application in airborne systems. To address this issue, a robustness-oriented design method for airborne visual perception based on deep ensemble learning is established. First, a highly representative dataset is generated based on the operational design domain, and a K-fold cross-validation method based on CW-SSIM is proposed to improve the independence between the training and validation sets with limited data. Second, based on the YOLO architecture, depthwise separable convolution is introduced, and three optimized base learners are designed to address different detection needs through multi-scale feature fusion, enhanced focus on small object detection, and fine-grained feature extraction. Finally, an ensemble learning method is designed using a weighted adaptive fusion strategy to dynamically adjust the weights of base learners, thereby improving the model accuracy and robustness. Experimental results show that the ensemble learning model outperforms detection box fusion algorithms such as NMS and WBF. When the IoU is not less than 0.7, the ensemble model improves the average P-value, R-value, and F1 score by at least 11.36%, 2.06%, and 6.78%, respectively, compared to a single model. When the IoU is no less than 0.75, the AP value increases by at least approximately 3%. These results indicate that proposed method significantly enhances target detection accuracy and robustness in complex environments, effectively reducing false positives and missed detections, and provides technical assurance for the safe flight of aircraft.

Cite this article

Zan MA , Tongjie ZHANG , Jie BAI , Yong CHEN , Yi TIAN . Robustness design and analysis of airborne visual perception based on deep ensemble learning[J]. ACTA AERONAUTICAET ASTRONAUTICA SINICA, 2026 , 47(12) : 332898 -332898 . DOI: 10.7527/S1000-6893.2025.32898

References

[1] 张辉, 杜瑞, 钟杭, 等. 电力设施多模态精细化机器人巡检关键技术及应用[J]. 自动化学报202551(1): 20-42.
  ZHANG H, DU R, ZHONG H, et al. The key technology and application of multi-modal fine robot inspection for power facilities[J]. Acta Automatica Sinica202551(1): 20-42 (in Chinese).
[2] 肖扬, 周军. 图像边缘检测综述[J]. 计算机工程与应用202359(5): 40-54.
  XIAO Y, ZHOU J. Overview of image edge detection[J]. Computer Engineering and Applications202359(5): 40-54 (in Chinese).
[3] 宫金良, 孙科, 张彦斐, 等. 基于梯度下降和角点检测的玉米根茎定位导航线提取方法[J]. 农业工程学报202238(13): 177-183.
  GONG J L, SUN K, ZHANG Y F, et al. Extracting navigation line for rhizome location using gradient descent and corner detection[J]. Transactions of the Chinese Society of Agricultural Engineering202238(13): 177-183 (in Chinese).
[4] LI Y H, GAN X C. An integrated fast Hough transform for multidimensional data[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence202345(9): 11365-11373.
[5] LIU Z X, SHAO F. Feature point matching based on multi-scale local relative motion consistency[J]. IEEE Access202311: 124845.
[6] 王秋富, 石治国, 张倬, 等. 舰载机着舰引导中鲁棒单目视觉相对位姿测量[J]. 航空学报202445(23): 330309.
  WANG Q F, SHI Z G, ZHANG Z, et al. Robust monocular relative pose measurement for carrier-based aircraft landing guidance[J]. Acta Aeronautica et Astronautica Sinica202445(23): 330309 (in Chinese).
[7] 陈树生, 贾苜梁, 林家豪, 等. 生成式模型赋能飞行器技术应用研究进展与展望[J]. 航空学报202546(10): 631194.
  CHEN S S, JIA M L, LIN J H, et al. Empowering aircraft technology applications with generative models: Research progress and prospects[J]. Acta Aeronautica et Astronautica Sinica202546(10): 631194 (in Chinese).
[8] BALDUZZI G, FERRARI B, CHERNOVA A, et al. Neural network based runway landing guidance for general aviation autoland: DOT/FAA/TC-21/48[R]. Atlantic: FAA, 2021.
[9] 马赞, 白杰, 陈勇, 等. 基于条件高斯PAC-Bayes的机载CNN分类器安全性评估[J]. 航空学报202546(4): 330824.
  MA Z, BAI J, CHEN Y, et al. Safety assessment for airborne CNN classifier based on conditional Gaussian PAC-Bayes[J]. Acta Aeronautica et Astronautica Sinica202546(4): 330824 (in Chinese).
[10] 倪静, 马波, 杨朝旭, 等. 视觉/惯性着陆组合引导方法设计与试验[J]. 航空学报202344(S1): 727636.
  NI J, MA B, YANG Z X, et al. Design and test of visual-inertial integrated method for landing guidance[J].Acta Aeronautica et Astronautica Sinica202344(S1): 727636 (in Chinese).
[11] 姜凌峰, 李新凯, 张海, 等. 基于改进 TD3 算法的无人机动态环境无地图导航[J]. 航空学报202546(8): 331035.
  JIANG L F, LI X K, ZHANG H, et al. Mapless navigation of UAVs in dynamic environments based on an improved TD3 algorithm[J]. Acta Aeronautica et Astronautica Sinica202546(8): 331035 (in Chinese).
[12] U.S. Department of Defense. Unmanned systems integrated roadmap: Fiscal years 2017-2042[R]. Washington, D.C.: Office of the Secretary of Defense, 2017.
[13] MICLEA V C, NEDEVSCHI S. Monocular depth estimation with improved long-range accuracy for UAV environment perception[J]. IEEE Transactions on Geoscience and Remote Sensing202260: 5602215.
[14] WANG C Y, MENG L L, GAO Q, et al. A target sensing and visual tracking method for countering unmanned aerial vehicle swarm[J]. IEEE Sensors Journal202424(19): 30340-30351.
[15] WANG S, CLARK R, WEN H K, et al. DeepVO: Towards end-to-end visual odometry with deep recurrent convolutional neural networks[C]∥2017 IEEE International Conference on Robotics and Automation (ICRA). Piscataway: IEEE Press, 2017: 2043-2050.
[16] ROBERTS D R, BAHN V, CIUTI S, et al. Cross-validation strategies for data with temporal, spatial, hierarchical, or phylogenetic structure[J]. Ecography201740(8): 913-929.
[17] YUE P Y, XIN J, ZHANG Y M, et al. Semantic-driven autonomous visual navigation for unmanned aerial vehicles[J]. IEEE Transactions on Industrial Electronics202471(11): 14853-14863.
[18] SOLOVYEV R, WANG W M, GABRUSEVA T. Weighted boxes fusion: Ensembling boxes from different object detection models[J]. Image and Vision Computing2021107: 104117.
[19] ZHOU H J, LI Z C, NING C C, et al. CAD: Scale invariant framework for real-time object detection[C]∥ 2017 IEEE International Conference on Computer Vision Workshops (ICCVW). Piscataway: IEEE Press, 2018: 760-768.
[20] YASEEN M. What is YOLOv8: An in-depth exploration of the internal features of the next-generation object detector[DB/OL]. arXiv preprint: 2408.15857, 2024.
[21] HOWARD A G, ZHU M L, CHEN B, et al. MobileNets: Efficient convolutional neural networks for mobile vision applications[DB/OL]. arXiv preprint: 1704. 04861, 2017.
[22] JEAN C, XAVIER H. Concepts of design assurance for neural networks [R]. Cologne: EASA, 2020.
[23] WANG Z, BOVIK A C, SHEIKH H R, et al. Image quality assessment: From error visibility to structural similarity[J]. IEEE Transactions on Image Processing200413(4): 600-612.
[24] SAMPAT M P, WANG Z, GUPTA S, et al. Complex wavelet structural similarity: A new image similarity index[J]. IEEE Transactions on Image Processing200918(11): 2385-2401.
[25] GE H M, WANG L G, PAN H Z, et al. Affinity propagation based on structural similarity index and local outlier factor for hyperspectral image clustering[J]. Remote Sensing202214(5): 1195.
[26] FIGUEIREDO R B D, MENDES H A. Analyzing information leakage on video object detection datasets by splitting images into clusters with high spatiotemporal correlation[J]. IEEE Access202412: 47646-47655.
[27] RADFORD A, KIM J W, HALLACY C, et al. Learning transferable visual models from natural language supervision[C]∥Proceedings of the 38th International Conference on Machine Learning (ICML 2021). Proceedings of Machine Learning Research, 2021139: 8748-8763.
[28] DENG D S. DBSCAN clustering algorithm based on density[C]∥2020 7th International Forum on Electrical Engineering and Automation (IFEEA). Piscataway: IEEE Press, 2021: 949-953.
[29] ANKERST M, BREUNIG M M, KRIEGEL H P, et al. OPTICS: Ordering points to identify the clustering structure[J]. ACM SIGMOD Record199928(2): 49-60.
[30] TOKUDA E K, COMIN C H, COSTA L DA F. Revisiting agglomerative clustering[J]. Physica A: Statistical Mechanics and Its Applications2022585: 126433.
Outlines

/