• 文献检索
  • 文档翻译
  • 深度研究
  • 学术资讯
  • Suppr Zotero 插件Zotero 插件
  • 邀请有礼
  • 套餐&价格
  • 历史记录
应用&插件
Suppr Zotero 插件Zotero 插件浏览器插件Mac 客户端Windows 客户端微信小程序
定价
高级版会员购买积分包购买API积分包
服务
文献检索文档翻译深度研究API 文档MCP 服务
关于我们
关于 Suppr公司介绍联系我们用户协议隐私条款
关注我们

Suppr 超能文献

核心技术专利:CN118964589B侵权必究
粤ICP备2023148730 号-1Suppr @ 2026

文献检索

告别复杂PubMed语法,用中文像聊天一样搜索,搜遍4000万医学文献。AI智能推荐,让科研检索更轻松。

立即免费搜索

文件翻译

保留排版,准确专业,支持PDF/Word/PPT等文件格式,支持 12+语言互译。

免费翻译文档

深度研究

AI帮你快速写综述,25分钟生成高质量综述,智能提取关键信息,辅助科研写作。

立即免费体验

利用深度卷积神经场从单目图像中学习深度。

Learning Depth from Single Monocular Images Using Deep Convolutional Neural Fields.

出版信息

IEEE Trans Pattern Anal Mach Intell. 2016 Oct;38(10):2024-39. doi: 10.1109/TPAMI.2015.2505283. Epub 2015 Dec 3.

DOI:10.1109/TPAMI.2015.2505283
PMID:26660697
Abstract

In this article, we tackle the problem of depth estimation from single monocular images. Compared with depth estimation using multiple images such as stereo depth perception, depth from monocular images is much more challenging. Prior work typically focuses on exploiting geometric priors or additional sources of information, most using hand-crafted features. Recently, there is mounting evidence that features from deep convolutional neural networks (CNN) set new records for various vision applications. On the other hand, considering the continuous characteristic of the depth values, depth estimation can be naturally formulated as a continuous conditional random field (CRF) learning problem. Therefore, here we present a deep convolutional neural field model for estimating depths from single monocular images, aiming to jointly explore the capacity of deep CNN and continuous CRF. In particular, we propose a deep structured learning scheme which learns the unary and pairwise potentials of continuous CRF in a unified deep CNN framework. We then further propose an equally effective model based on fully convolutional networks and a novel superpixel pooling method, which is about 10 times faster, to speedup the patch-wise convolutions in the deep model. With this more efficient model, we are able to design deeper networks to pursue better performance. Our proposed method can be used for depth estimation of general scenes with no geometric priors nor any extra information injected. In our case, the integral of the partition function can be calculated in a closed form such that we can exactly solve the log-likelihood maximization. Moreover, solving the inference problem for predicting depths of a test image is highly efficient as closed-form solutions exist. Experiments on both indoor and outdoor scene datasets demonstrate that the proposed method outperforms state-of-the-art depth estimation approaches.

摘要

在本文中,我们解决了从单目图像进行深度估计的问题。与使用多个图像(如立体深度感知)进行深度估计相比,从单目图像进行深度估计更具挑战性。之前的工作通常侧重于利用几何先验或其他信息源,大多数使用手工制作的特征。最近,越来越多的证据表明,来自深度卷积神经网络(CNN)的特征为各种视觉应用创造了新的记录。另一方面,考虑到深度值的连续性,深度估计可以自然地表述为连续条件随机场(CRF)学习问题。因此,这里我们提出了一种从单目图像估计深度的深度卷积神经场模型,旨在共同探索深度 CNN 和连续 CRF 的能力。具体来说,我们提出了一种深度结构学习方案,在统一的深度 CNN 框架中学习连续 CRF 的一元和二元势。然后,我们进一步提出了一种基于全卷积网络和新颖的超像素池化方法的等效有效模型,该模型的速度大约快 10 倍,以加速深度模型中的补丁卷积。通过这个更有效的模型,我们能够设计更深的网络以追求更好的性能。我们提出的方法可以用于没有几何先验或任何额外信息注入的一般场景的深度估计。在我们的案例中,可以通过闭式形式计算配分函数的积分,从而可以精确地求解对数似然最大化。此外,由于存在闭式解,预测测试图像深度的推断问题的效率非常高。在室内和室外场景数据集上的实验表明,所提出的方法优于最新的深度估计方法。

相似文献

1
Learning Depth from Single Monocular Images Using Deep Convolutional Neural Fields.利用深度卷积神经场从单目图像中学习深度。
IEEE Trans Pattern Anal Mach Intell. 2016 Oct;38(10):2024-39. doi: 10.1109/TPAMI.2015.2505283. Epub 2015 Dec 3.
2
Deep Learning-Based Monocular Depth Estimation Methods-A State-of-the-Art Review.基于深度学习的单目深度估计方法——最新综述。
Sensors (Basel). 2020 Apr 16;20(8):2272. doi: 10.3390/s20082272.
3
Monocular Depth Estimation Using Multi-Scale Continuous CRFs as Sequential Deep Networks.使用多尺度连续条件随机场作为序列深度网络的单目深度估计
IEEE Trans Pattern Anal Mach Intell. 2019 Jun;41(6):1426-1440. doi: 10.1109/TPAMI.2018.2839602. Epub 2018 May 22.
4
Discriminative Training of Deep Fully Connected Continuous CRFs With Task-Specific Loss.基于任务特定损失的深度全连接连续条件随机场的判别式训练
IEEE Trans Image Process. 2017 May;26(5):2127-2136. doi: 10.1109/TIP.2017.2675166. Epub 2017 Feb 24.
5
Deep learning and conditional random fields-based depth estimation and topographical reconstruction from conventional endoscopy.基于深度学习和条件随机场的传统内窥镜深度估计和地形重建。
Med Image Anal. 2018 Aug;48:230-243. doi: 10.1016/j.media.2018.06.005. Epub 2018 Jun 14.
6
Improving dense conditional random field for retinal vessel segmentation by discriminative feature learning and thin-vessel enhancement.通过判别特征学习和细血管增强改进用于视网膜血管分割的密集条件随机场
Comput Methods Programs Biomed. 2017 Sep;148:13-25. doi: 10.1016/j.cmpb.2017.06.016. Epub 2017 Jun 24.
7
A novel end-to-end classifier using domain transferred deep convolutional neural networks for biomedical images.一种使用域转移深度卷积神经网络的新型端到端生物医学图像分类器。
Comput Methods Programs Biomed. 2017 Mar;140:283-293. doi: 10.1016/j.cmpb.2016.12.019. Epub 2017 Jan 6.
8
MR-based synthetic CT generation using a deep convolutional neural network method.基于磁共振成像利用深度卷积神经网络方法生成合成CT图像
Med Phys. 2017 Apr;44(4):1408-1419. doi: 10.1002/mp.12155. Epub 2017 Mar 21.
9
Classification of CT brain images based on deep learning networks.基于深度学习网络的 CT 脑图像分类。
Comput Methods Programs Biomed. 2017 Jan;138:49-56. doi: 10.1016/j.cmpb.2016.10.007. Epub 2016 Oct 20.
10
Benchmark Data Set and Method for Depth Estimation from Light Field Images.用于光场图像深度估计的基准数据集和方法。
IEEE Trans Image Process. 2018 Jul;27(7):3586-3598. doi: 10.1109/TIP.2018.2814217. Epub 2018 Mar 9.

引用本文的文献

1
Vision-Based Collision Warning Systems with Deep Learning: A Systematic Review.基于深度学习的视觉碰撞预警系统:系统综述
J Imaging. 2025 Feb 17;11(2):64. doi: 10.3390/jimaging11020064.
2
Camera-view supervision for bird's-eye-view semantic segmentation.用于鸟瞰语义分割的相机视图监督
Front Big Data. 2024 Nov 15;7:1431346. doi: 10.3389/fdata.2024.1431346. eCollection 2024.
3
LapUNet: a novel approach to monocular depth estimation using dynamic laplacian residual U-shape networks.LapUNet:一种使用动态拉普拉斯残差U型网络进行单目深度估计的新方法。
Sci Rep. 2024 Oct 9;14(1):23544. doi: 10.1038/s41598-024-74445-x.
4
Three-Dimensional Dense Reconstruction: A Review of Algorithms and Datasets.三维密集重建:算法与数据集综述
Sensors (Basel). 2024 Sep 10;24(18):5861. doi: 10.3390/s24185861.
5
Monocular Depth Estimation via Self-Supervised Self-Distillation.通过自监督自蒸馏进行单目深度估计
Sensors (Basel). 2024 Jun 24;24(13):4090. doi: 10.3390/s24134090.
6
Study on Gesture Recognition Method with Two-Stream Residual Network Fusing sEMG Signals and Acceleration Signals.基于双通道残差网络融合表面肌电信号和加速度信号的手势识别方法研究。
Sensors (Basel). 2024 Apr 24;24(9):2702. doi: 10.3390/s24092702.
7
Multi-Spectral Food Classification and Caloric Estimation Using Predicted Images.利用预测图像进行多光谱食物分类与热量估计
Foods. 2024 Feb 11;13(4):551. doi: 10.3390/foods13040551.
8
Uncovering local aggregated air quality index with smartphone captured images leveraging efficient deep convolutional neural network.利用高效深度卷积神经网络,通过智能手机拍摄的图像揭示局部空气质量综合指数。
Sci Rep. 2024 Jan 18;14(1):1627. doi: 10.1038/s41598-023-51015-1.
9
Unsupervised Monocular Depth and Camera Pose Estimation with Multiple Masks and Geometric Consistency Constraints.基于多掩码和几何一致性约束的无监督单目深度与相机位姿估计
Sensors (Basel). 2023 Jun 4;23(11):5329. doi: 10.3390/s23115329.
10
Monocular metasurface camera for passive single-shot 4D imaging.用于被动式单次 4D 成像的单目超表面相机。
Nat Commun. 2023 Feb 23;14(1):1035. doi: 10.1038/s41467-023-36812-6.