| [1]ZHAN Y, XIONG Z, YUAN Y.RSVG: Exploring data and models for visual grounding on remote sensing data[J]. IEEE Transactions on Geoscience and Remote Sensing, 2023, 61: 1-13.[J].IEEE Transactions on Geoscience and Remote Sensing, 2023, 61(无):1-13
[2]SUN Y, FENG S, LI X, et al.Visual grounding in remote sensing images[C]//Proceedings of the ACM International Conference on Multimedia. New York: ACM, 2022: 404-412..Proceedings of the ACM International Conference on Multimedia, 2022, 无(无):404-412
[3]LI T, ZHANG Y, WANG C, et al.TACMT: Text-aware cross-modal transformer for visual grounding on high-resolution SAR images[J]. ISPRS Journal of Photogrammetry and Remote Sensing, 2025, 222: 152-166.[J].ISPRS Journal of Photogrammetry and Remote Sensing, 2025, 222(无):152-166
[4]MAO J, HUANG J, TOSHEV A, et al.Generation and comprehension of unambiguous object descriptions[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway: IEEE, 2016: 11-20..Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2019, 无(无):11-20
[5]YANG Z, GONG B, WANG L, et al.A joint speaker-listener-reinforcer model for referring expression comprehension[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway: IEEE, 2019: 12404-12413..Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2019, 无(无):12404-12413
[6]DENG J, YANG Z, CHEN T, et al.TransVG: End-to-end visual grounding with transformers[C]//Proceedings of the IEEE/CVF International Conference on Computer Vision. Piscataway: IEEE, 2021: 1769-1779..Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021, 无(无):1769-1779
[7]YANG L, XU Y, YUAN C, et al.Improving visual grounding with visual-linguistic verification and iterative reasoning[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway: IEEE, 2022: 9499-9508..Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022, 无(无):9499-9508
[8]WANG F, WU C, WU J, et al.Multistage synergistic aggregation network for remote sensing visual grounding[J]. IEEE Geoscience and Remote Sensing Letters, 2024, 21: 1-5.[J].IEEE Geoscience and Remote Sensing Letters, 2024, 21(无):1-5
[9]DING Y, WANG D, LI K, et al.Visual grounding of remote sensing images with multi-dimensional semantic-guidance[J]. Pattern Recognition Letters, 2025, 189: 85-91.[J].Pattern Recognition Letters, 2025, 189(无):85-91
[10]MA Q, PAN J, BAI C.Direction-oriented visual-semantic embedding model for remote sensing image-text retrieval[J]. IEEE Transactions on Geoscience and Remote Sensing, 2024, 62: 4704014.[J].IEEE Transactions on Geoscience and Remote Sensing, 2024, 62(无):1-14
[11]WANG M, GUO J, SONG B, et al.Graph-based hierarchical semantic consistency network for remote sensing image-text retrieval[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2025, 18: 15334-15346.[J].IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2025, 18(无):15334-15346
[12]LI J, ZHANG Y, WANG C, et al.Multi-scale cross-modal attention fusion for remote sensing visual grounding[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2024, 17: 3215-3228.[J].IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2024, 17(无):3215-3228
[13]YUAN Z, ZHANG W, RONG X, et al.A lightweight multi-scale crossmodal text-image retrieval method in remote sensing[J]. IEEE Transactions on Geoscience and Remote Sensing, 2021, 60: 1-19.[J].IEEE Transactions on Geoscience and Remote Sensing, 2021, 60(无):1-19
[14]CHENG Q, ZHOU Y, FU P, et al.A deep semantic alignment network for the cross-modal image-text retrieval in remote sensing[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2021, 14: 4284-4297.[J].IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2021, 14(无):4284-4297
[15]YU H, YAO F, LU W, et al.Text-image matching for cross-modal remote sensing image retrieval via graph neural network[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2022, 16: 812-824.[J].IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2022, 16(无):812-824
[16]FU K, ZHANG X, LIU G, et al.Scattering-keypoint-guided network for oriented ship detection in high-resolution and large-scale SAR images[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2021, 14: 11162-11178.[J].IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2021, 14(无):11162-11178
[17]LU X, WANG B, ZHENG X, et al.Exploring models and data for remote sensing image caption generation[J].IEEE Transactions on Geoscience and Remote Sensing, 2017, 56(4):2183-2195
[18]LI Y, MAO H, GIRSHICK R, et al.Exploring plain vision transformer backbones for object detection[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway: IEEE, 2022: 2804-2814..Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2022, 无(无):2804-2814
[19]LIU Y, OTT M, GOYAL N, et al.RoBERTa: A robustly optimized BERT pretraining approach[EB/OL]. (2019-07-26)[2026-03-30]. https://arxiv.org/abs/1907.11692.[J].arXiv preprint arXiv:1907.11692, 2019, 无(无):无-无
[20]REN S, HE K, GIRSHICK R, et al.Faster R-CNN: Towards real-time object detection with region proposal networks[J].IEEE Transactions on Pattern Analysis and Machine Intelligence, 2017, 39(6):1137-1149
[21]YANG Z, GONG B, WANG L, et al.A fast and accurate one-stage approach to visual grounding[C]//Proceedings of the IEEE/CVF International Conference on Computer Vision. Piscataway: IEEE, 2019: 4683-4693..Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2019, 无(无):4683-4693
[22]YANG Z, CHEN T, WANG L, et al.Improving one-stage visual grounding by recursive sub-query construction[C]//Proceedings of the European Conference on Computer Vision. Cham: Springer, 2020: 387-404..Proceedings of the European Conference on Computer Vision (ECCV), 2020, 无(无):387-404
[23]HUANG B, LIAN D, LUO W, et al.Look before you leap: learning landmark features for one-stage visual grounding[C]//Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Piscataway: IEEE, 2021: 16888-16897..Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2021, 无(无):16888-16897 |