基于注意力增强的跨模态多层融合网络的物体位姿估计
杨恒,王韶涵,董青,赵科渊,杨明亮

Object pose estimation based on attention enhancement cross-modal multilayer fusion network
Heng YANG,Shaohan WANG,Qing DONG,Keyuan ZHAO,Mingliang YANG
图 4 提取的多尺度密集判别特征中每个模块的输出大小
Fig.4 Output size of each module in extracted multi-scale dense discriminative features