[1] LI L, LI X J, YANG S L, et al.Unsupervised-learning-based continuous depth and motion estimation with monocular endoscopy for virtual reality minimally invasive surgery[J].IEEE transactions on industrial informatics, 2021, 17(6):3920-3928. [2] EIGEN D, PUHRSCH C, FERGUS R.Depth map prediction from a single image using a multi-scale deep network[A].Advances in neural information processing systems 27 (NIPS 2014)[C].Montreal, Canada, 2014.2366-2374. [3] PENG R, WANG R G, LAI Y W, et al.Excavating the potential capacity of self-supervised monocular depth estimation[A].2021 IEEE/CVF International Conference on Computer Vision (ICCV)[C].Montreal, QC, Canada:IEEE,2021.15270-15279. [4] VASWANI A, SHAZEER N, PARMAR N,et al.Attention is all you need[J].Advances in neural information processing systems, 2017, 30: 5998-6008. [5] SCHÖN M, BUCHHOLZ M, DIETMAYER K. MGNet: Monocular geometric scene understanding for autonomous driving[A].2021 IEEE/CVF International Conference on Computer Vision (ICCV)[C].Montreal, QC, Canada:IEEE, 2021.15511-15520. [6] GODARD C, MAC AODHA O, BROSTOW G J.Unsupervised monocular depth estimation with left-right consistency[A].2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)[C].Honolulu, HI, USA: IEEE,2017. 6602-6610. [7] ZHOU T H, BROWN M, SNAVELY N,et al. Unsupervised learning of depth and ego-motion from video[EB/OL].2017. https://arxiv.org/abs/1704.07813. [8] 严亚滔,卜鹏辉,王航,等.基于全局注意力机制的单目深度估计算法[J].计算机技术与发展,2025,35(6):34-41. [9] 孙琦,胡建华,胡立华,等.融合转置注意力的自监督单目深度估计方法[J].计算机技术与发展, 2025 ,35(9) : 38-45. [10] 宋磊, 李嵘, 焦义涛, 等. 基于ResNeXt单目深度估计的幼苗植株高度测量方法[J].农业工程学报, 2022, 38(3):155-163. [11] 周云成, 邓寒冰, 许童羽, 等. 基于稠密自编码器的无监督番茄植株图像深度估计模型[J].农业工程学报, 2020, 36(11):182-192. [12] CUI X Z, FENG Q, WANG S Z, et al.Monocular depth estimation with self-supervised learning for vineyard unmanned agricultural vehicle[J].Sensors, 2022, 22(3):721. [13] UHRIG J, SCHNEIDER N, SCHNEIDER L, et al. Sparsity invariant CNNs[EB/OL].2017. https://arxiv.org/abs/1708.06500. [14] GAO X, LI J, WANG Y.MViTDepth: Efficient Monocular Depth Estimation with MobileViT[J].IEEE Transactions on intelligent transportation systems, 2023, 24: 8210-8221. [15] JOHNSTON A,CARNEIRO G. Self-supervised monocular trained depth estimation using self-attention and discrete disparity volume[EB/OL].2020. https://arxiv.org/abs/2003.13951. [16] LIU L N, SONG X B, WANG M M, et al.Self-supervised monocular depth estimation for all day images using domain separation[A].2021 IEEE/CVF International Conference on Computer Vision[C].Montreal, QC, Canada:IEEE,2021.12717-12726. [17] POGGI M, ALEOTTI F, TOSI F, et al.On the uncertainty of self-supervised monocular depth estimation[A].2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)[C].Seattle, WA, USA:IEEE,2020.3224-3234. [18] KINGMA D P, BA J. Adam: A method for stochastic optimization[EB/OL].2014. https://arxiv.org/abs/1412.6980. [19] MEHTA S, RASTEGARI M. MobileViT: Light-weight, General-purpose,Mobile-friendly Vision Transformer[EB/OL].2021. https://arxiv.org/abs/2110.02178. [20] XUE F, ZHUO G R, HUANG Z Y, et al.Toward hierarchical self-supervised monocular absolute depth estimation for autonomous driving applications [A].2020 IEEE/RSJ International Conference on Intelligent Robots and Systems[C].Las Vegas, NV, USA:IEEE, 2020.2330-2337. [21] GODARD C, MAC AODHA O, FIRMAN M,et al.Digging into self-supervised monocular depth estimation [A].IEEE/CVF International conference on computer vision (ICCV)[C].Seoul, Korea (South):IEEE, 2019. 3828-3838. [22] XIANG J, WANG Y, AN L F, et al.Visual attention-based self-supervised absolute depth estimation using geometric priors in autonomous driving[J].IEEE robotics and automation letters, 2022, 7(4):11998-12005. [23] HE K M, ZHANG X Y, REN S Q,et al.Deep residual learning for image recognition [A].2016 IEEE Conference on computer vision and pattern recognition (CVPR)[C].Las Vegas, NV, USA:IEEE,2016. 770-778. [24] DOSOVITSKIY A, BEYER L, KOLESNIKOV A,et al. An Image is Worth 16x16 Words: Transformers for image recognition at scale [EB/OL].2021.https://arxiv.org/abs/2010.11929. [25] ZHANG H, CHEN Y, LIU Z.EMO: Efficient mobile model for dense prediction tasks[J].Advances in neural information processing systems, 2023, 36: 12345-12356. [26] ZHANG N, NEX F, VOSSELMAN G,et al. Lite-Mono: A lightweight CNN and transformer architecture for self-supervised monocular depth estimation [EB/OL].2022. https://doi.org/10.48550/arXiv.2211.13202. [27] 张晖敏. 基于特征融合和垂直敏感性的单目自监督深度估计研究[D].江苏徐州:中国矿业大学,2024. [28] 贾迪,宋慧伦,赵辰,等.面向单目深度估计的多层次感知条件随机场模型[J].中国图象图形学报,2025,30(3):824-841. [29] HINTON G, VINYALS O, DEAN J. Distilling the knowledge in a neural network [EB/OL].2015. https://arxiv.org/abs/1503.02531. [30] SANDLER M, HOWARD A, ZHU M L,et al. MobileNetV2: Inverted residuals and linear bottlenecks[EB/OL].2018. https://arxiv.org/abs/1801.04381. |