This paper proposes a fusion algorithm utilizing TransNeXt to address detail loss and artifact generation issues in the fusion of infrared and visible images. Firstly, shallow and deep features are extracted from the source images using convolutional neural networks and TransNeXt. An information compensation module is employed to enhance the semantic information of the infrared shallow features. Secondly, a cross-attention-based fusion module integrates these features, and dynamically adjusts weights based on the importance of different regions in the source images to adapt to scene variations, thereby improving fusion robustness and accuracy. The final fused image is obtained through Transformer-based image reconstruction. In addition,the proposed method constrains the fusion process through a VGG19-based saliency mask loss function, preserving richer information in key regions of the fused results. The experimental results indicate that, compared with the other seven methods, this approach has improved the objective evaluation metrics: namely information entropy, standard deviation, sum of correlation differ-ences, peak signal-to-noise ratio, and pixel feature mutual information, by an average of 10.92%, 14.85%, 24.80%, 2.26%, and 1.30%, respectively. Furthermore, it effectively preserves rich texture information while minimizing artifacts, demonstrating outstanding performance in night light fusion. Additionally, it has achieved superior results in object detection relative to the comparison methods.
LEIJ, LIJ, LIUJ,et al .GALFusion:multi-exposure image fusion via a global-local aggregation learning network[J].IEEE Transactions on Instrumentation and Measurement,2023,72:1-15.
[2]
RAOD Y, XUT Y, WUX J .TGFuse:an infrared and visible image fusion approach based on transformer and generative adversarial network[J].IEEE Transactions on Image Processing,2023.
[3]
LIH, QIX, XIEW .Fast infrared and visible image fusion with structural decomposition[J]. Knowledge-Based Systems,2020,204:106182.
[4]
MAJ Y, MAY, LIC .Infrared and visible image fusion methods and applications:a survey[J]. Information Fusion,2019,45:153-178.
[5]
ZHANGH, XUH, TIANX,et al. Image fusion meets deep learning:a survey and perspective[J]. Information Fusion,2021,76:323-336.
[6]
MAJ Y, XUH, JIANGJ J,et al .DDcGAN:a dual-discriminator conditional generative adversarial network for multi-resolution image fusion[J]. IEEE Transactions on Image Processing,2020,29:4980-4995.
[7]
LIH, WUX J, KITTLERJ .RFN-Nest:an end-to-end residual fusion network for infrared and visible images[J]. Information Fusion, 2021, 73: 72-86.
[8]
ZHANGH, MAJ Y. SDNet:a versatile squeeze-and-decomposition network for real-time image fusion[J]. International Journal of Computer Vision,2021,129(10):2761-2785.
[9]
SHENS, LID, MEIL Y,et al .DFA-net:multi-scale dense feature-aware network via integrated attention for unmanned aerial vehicle infrared and visible image fusion[J].Drones,2023, 7(8): 517.
[10]
VASWANIA, SHAZEERN, PARMARN, et al. Attention is all you need [J]. Advances in Neural Information Processing Systems, 2017, 30(1): 261-272.
[11]
WANGZ S, CHENY L, SHAOW Y,et al .SwinFuse:a residual swin transformer fusion network for infrared and visible images[J].IEEE Transactions on Instrumentation and Measurement,2022,71:1-12.
[12]
LIUZ, HUH, LINY T,et al .Swin transformer V2:scaling up capacity and resolution[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.2022: 12009-12019.
[13]
TANGW, HEF Z, LIUY,et al .DATFuse:infrared and visible image fusion via dual attention transformer[J].IEEE Transactions on Circuits and Systems for Video Technology,2023,33(7):3159-3172.
[14]
DOSOVITSKIYA, BEYERL, KOLESNIKOVA, et al. An image is worth 16x16 words: Transformers for image recognition at scale[J/OL].arXiv preprint arXiv:2020.
[15]
SHID .TransNeXt:robust foveal visual perception for vision transformers[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.2024:17773-17783.
LIH, WUX J .CrossFuse:a novel cross attention mechanism based infrared and visible image fusion approach[J/OL].Information Fusion,2024,103:102147.
[18]
SIMONYANK, ZISSERMANA .Very deep convolutional networks for large-scale image recognition[J/OL]. arXiv preprint arXiv:2014.
[19]
LIUY, CHENX, WARDR K,et al .Image fusion with convolutional sparse representation[J].IEEE Signal Processing Letters,2016,23(12):1882-1886.
[20]
XUH, MAJ Y, JIANGJ J,et al. U2Fusion:a unified unsupervised image fusion network[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence,2020,44(1):502-518.
[21]
TANGL F, YUANJ T, MAJ Y .Image fusion in the loop of high-level vision tasks:a semantic-aware real-time infrared and visible image fusion network[J].Information Fusion,2022,82:28-42.
[22]
TANGL F, YUANJ T, ZHANGH,et al .PIAFusion:a progressive infrared and visible image fusion network based on illumination aware[J].Information Fusion,2022,83:79-92.
[23]
TANGW, HEF Z, LIUY,et al. DATFuse:infrared and visible image fusion via dual attention transformer[J]. IEEE Transactions on Circuits and Systems for Video Technology,2023,33(7):3159-3172.
[24]
ZHAOZ X, BAIH W, ZHANGJ S,et al .CDDFuse:correlation-driven dual-branch feature decomposition for multi-modality image fusion[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2023:5906-5916.
[25]
TANGL F, ZHANGH, XUH,et al .Rethinking the necessity of image fusion in high-level vision tasks:a practical infrared and visible image fusion network based on progressive semantic injection and scene fidelity[J]. Information Fusion,2023,99:101870.
[26]
REDMONJ, DIVVALAS, GIRSHICKR,et al. You only look once:unified,real-time object detection[C]//Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition.2016: 779-788.
基金资助
国家自然科学基金资助项目(62462043)
国家自然科学基金资助项目(62067006)
National Natural ScienceFoundation of China(62462043)
National Natural ScienceFoundation of China(62067006)
甘肃省重点研发计划项目(25YFGA047)
Key Research and Development Project of Gansu Province(25YFGA047)
甘肃省自然科学基金项目(23JRRA847)
Natural Science Foundation of Gansu Province(23JRRA847)