Deep Learning Advances in Computer Vision with 3D Data 论文

2017ACM Computing Surveys引用 328
Advanced Neural Network ApplicationsAdvanced Image and Video Retrieval Techniques3D Surveying and Cultural Heritage

摘要

Deep learning has recently gained popularity achieving state-of-the-art performance in tasks involving text, sound, or image processing. Due to its outstanding performance, there have been efforts to apply it in more challenging scenarios, for example, 3D data processing. This article surveys methods applying deep learning on 3D data and provides a classification based on how they exploit them. From the results of the examined works, we conclude that systems employing 2D views of 3D data typically surpass voxel-based (3D) deep models, which however, can perform better with more layers and severe data augmentation. Therefore, larger-scale datasets and increased resolutions are required.