| 计算机技术 |
|
|
|
|
| 结合边缘辅助与多级特征融合的跨模态语义分割算法 |
陈广秋( ),任天蓉,段锦*( ),黄丹丹 |
| 长春理工大学 电子信息工程学院,吉林 长春 130022 |
|
| Cross-modal semantic segmentation algorithm with edge-assisted and multi-level feature fusion |
Guangqiu CHEN( ),Tianrong REN,Jin DUAN*( ),Dandan HUANG |
| College of Electronic Information Engineering, Changchun University of Science and Technology, Changchun 130022, China |
引用本文:
陈广秋,任天蓉,段锦,黄丹丹. 结合边缘辅助与多级特征融合的跨模态语义分割算法[J]. 浙江大学学报(工学版), 2026, 60(8): 1782-1791.
Guangqiu CHEN,Tianrong REN,Jin DUAN,Dandan HUANG. Cross-modal semantic segmentation algorithm with edge-assisted and multi-level feature fusion. Journal of ZheJiang University (Engineering Science), 2026, 60(8): 1782-1791.
链接本文:
https://www.zjujournals.com/eng/CN/10.3785/j.issn.1008-973X.2026.08.017
或
https://www.zjujournals.com/eng/CN/Y2026/V60/I8/1782
|
| 1 |
ROMERA E, ÁLVAREZ J M, BERGASA L M, et al ERFNet: efficient residual factorized ConvNet for real-time semantic segmentation[J]. IEEE Transactions on Intelligent Transportation Systems, 2018, 19 (1): 263- 272
doi: 10.1109/TITS.2017.2750080
|
| 2 |
FAN X, WANG X, GAO J, et al. Bi-level learning of task-specific decoders for joint registration and one-shot medical image segmentation [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Seattle: IEEE, 2024: 11726–11735.
|
| 3 |
CHEN H, LUO H, WANG C AfaMamba: adaptive feature aggregation with visual state space model for remote sensing images semantic segmentation[J]. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2025, 18: 8965- 8983
doi: 10.1109/JSTARS.2025.3552942
|
| 4 |
张振利, 胡新凯, 李凡, 等 基于CNN和Efficient Transformer的多尺度遥感图像语义分割算法[J]. 浙江大学学报: 工学版, 2025, 59 (4): 778- 786 ZHANG Zhenli, HU Xinkai, LI Fan, et al Semantic segmentation algorithm for multiscale remote sensing images based on CNN and Efficient Transformer[J]. Journal of Zhejiang University: Engineering Science, 2025, 59 (4): 778- 786
doi: 10.3785/j.issn.1008-973X.2025.04.013
|
| 5 |
SHELHAMER E, LONG J, DARRELL T Fully convolutional networks for semantic segmentation[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2017, 39 (4): 640- 651
doi: 10.1109/TPAMI.2016.2572683
|
| 6 |
BADRINARAYANAN V, KENDALL A, CIPOLLA R SegNet: a deep convolutional encoder-decoder architecture for image segmentation[J]. IEEE Transactions on Pattern Analysis and Machine Intelligence, 2017, 39 (12): 2481- 2495
doi: 10.1109/TPAMI.2016.2644615
|
| 7 |
RONNEBERGER O, FISCHER P, BROX T. U-Net: convolutional networks for biomedical image segmentation [C]// Medical Image Computing and Computer-Assisted Intervention. [S.l.]: Springer, 2015: 234–241.
|
| 8 |
ZHAO H, SHI J, QI X, et al. Pyramid scene parsing network [C]// Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. Honolulu. IEEE, 2017: 6230–6239.
|
| 9 |
XIE E, WANG W, YU Z, et al. SegFormer: simple and efficient design for semantic segmentation with transformers [EB/OL]. (2021–10–28)[2026–04–28]. https://arxiv.org/pdf/2105.15203.
|
| 10 |
HA Q, WATANABE K, KARASAWA T, et al. MFNet: towards real-time semantic segmentation for autonomous vehicles with multi-spectral scenes [C]// Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems. Vancouver: IEEE, 2017: 5108–5115.
|
| 11 |
SUN Y, ZUO W, LIU M RTFNet: RGB-thermal fusion network for semantic segmentation of urban scenes[J]. IEEE Robotics and Automation Letters, 2019, 4 (3): 2576- 2583
doi: 10.1109/LRA.2019.2904733
|
| 12 |
SHIVAKUMAR S S, RODRIGUES N, ZHOU A, et al. PST900: RGB-thermal calibration, dataset and segmentation network [C]// Proceedings of the IEEE International Conference on Robotics and Automation. Paris: IEEE, 2020: 9441–9447.
|
| 13 |
DENG F, FENG H, LIANG M, et al. FEANet: feature-enhanced attention network for RGB-thermal real-time semantic segmentation [C]// Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems. Prague: IEEE, 2021: 4467–4473.
|
| 14 |
ZHANG Q, ZHAO S, LUO Y, et al. ABMDRNet: adaptive-weighted bi-directional modality difference reduction network for RGB-T semantic segmentation [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Nashville: IEEE, 2021: 2633–2642.
|
| 15 |
YI S, CHEN M, LIU X, et al HAFFseg: RGB-Thermal semantic segmentation network with hybrid adaptive feature fusion strategy[J]. Signal Processing: Image Communication, 2023, 117: 117027
doi: 10.1016/j.image.2023.117027
|
| 16 |
ZHOU W, DONG S, FANG M, et al CACFNet: cross-modal attention cascaded fusion network for RGB-T urban scene parsing[J]. IEEE Transactions on Intelligent Vehicles, 2024, 9 (1): 1919- 1929
doi: 10.1109/TIV.2023.3314527
|
| 17 |
HE K, ZHANG X, REN S, et al. Deep residual learning for image recognition [C]// Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. Las Vegas: IEEE, 2016: 770–778.
|
| 18 |
BERMAN M, TRIKI A R, BLASCHKO M B. The lovasz-softmax loss: a tractable surrogate for the optimization of the intersection-over-union measure in neural networks [C]// Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. Salt Lake City: IEEE, 2018: 4413–4421.
|
| 19 |
SUN Y, ZUO W, YUN P, et al FuseSeg: semantic segmentation of urban scenes based on RGB and thermal data fusion[J]. IEEE Transactions on Automation Science and Engineering, 2021, 18 (3): 1000- 1011
doi: 10.1109/TASE.2020.2993143
|
| 20 |
HE X, WANG M, LIU T, et al SFAF-MA: spatial feature aggregation and fusion with modality adaptation for RGB-thermal semantic segmentation[J]. IEEE Transactions on Instrumentation and Measurement, 2023, 72: 5012810
doi: 10.1109/tim.2023.3267529
|
| 21 |
ZHOU W, DONG S, XU C, et al Edge-aware guidance fusion network for RGB–thermal scene parsing[J]. Proceedings of the AAAI Conference on Artificial Intelligence, 2022, 36 (3): 3571- 3579
doi: 10.1609/aaai.v36i3.20269
|
| 22 |
ZHOU Z, WU S, ZHU G, et al. Channel and spatial relation-propagation network for RGB-thermal semantic segmentation [EB/OL]. (2023–08–24)[2026–04–28]. https://arxiv.org/pdf/2308.12534.
|
|
Viewed |
|
|
|
Full text
|
|
|
|
|
Abstract
|
|
|
|
|
Cited |
|
|
|
|
| |
Shared |
|
|
|
|
| |
Discussed |
|
|
|
|