文献详情 >Advancing Underwater Vision: A... 收藏

Advancing Underwater Vision: A Survey of Deep Learning Models for Underwater Object Recognition and Tracking

作者：Elmezain, Mahmoud Saad Saoud, Lyes Sultan, Atif Heshmat, Mohamed Seneviratne, Lakmal Hussain, Irfan

作者机构：Khalifa Univ Khalifa Univ Ctr Autonomous Robot Syst Abu Dhabi U Arab Emirates

出版物：《IEEE ACCESS》 (IEEE Access)

年卷期：2025年第13卷

页面：17830-17867页

核心收录：

基　　金：Khalifa University of Science and Technology [8434000534, CIRA-2021-085, RC1-2018-KUCARS] KU-Stanford

主　　题：Reviews Object recognition Computer vision Computational modeling Imaging Image segmentation Deep learning Surveys Optical imaging Oceans Underwater computer vision deep learning underwater robotics ocean research underwater image enhancement object tracking object detection

摘要：Underwater computer vision plays a vital role in ocean research, enabling autonomous navigation, infrastructure inspections, and marine life monitoring. However, the underwater environment presents unique challenges, including color distortion, limited visibility, and dynamic light conditions, which hinder the performance of traditional image processing methods. Recent advancements in deep learning (DL) have demonstrated remarkable success in overcoming these challenges by enabling robust feature extraction, image enhancement, and object recognition. This review provides a comprehensive analysis of cutting-edge deep learning architectures designed for underwater object detection, segmentation, and tracking. State-of-the-art (SOTA) models, including AGW-YOLOv8, Feature-Adaptive FPN, and Dual-SAM, have shown substantial improvements in addressing occlusions, camouflaging, and small underwater object detection. For tracking tasks, transformer-based models like SiamFCA and FishTrack leverage hierarchical attention mechanisms and convolutional neural networks (CNNs) to achieve high accuracy and robustness in dynamic underwater environments. Beyond optical imaging, this review explores alternative modalities such as sonar, hyperspectral imaging, and event-based vision, which provide complementary data to enhance underwater vision systems. These approaches improve performance under challenging conditions, enabling richer and more informative scene interpretation. Promising future directions are also discussed, emphasizing the need for domain adaptation techniques to improve generalizability, lightweight architectures for real-time performance, and multi-modal data fusion to enhance interpretability and robustness. By critically evaluating current methodologies and highlighting gaps, this review provides insights for advancing underwater computer vision systems to support ocean exploration, ecological conservation, and disaster management.

本地馆藏 | 借阅须知 | 我要预约

已订购，未入库

sda

目录详情 | 试阅读 |

读者评论与其他读者分享你的观点

学校读者

用户名:未登录

我的评分

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

时间限定

文献类型

馆藏选择

核心期刊

语言

文献类型

帮助

文字说明：

检索规则说明：

检索范例：

分类表

所选分类

看过本文的还看了

相关文献

该作者的其他文献

CADAL相关文献

Advancing Underwater Vision: A Survey of Deep Learning Models for Underwater Object Recognition and Tracking

读者评论与其他读者分享你的观点

请选择收藏分类：

建议与咨询 留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

时间限定

文献类型

馆藏选择

核心期刊

语言

文献类型

帮助

文字说明：

检索规则说明：

检索范例：

分类表

所选分类

看过本文的还看了

相关文献

该作者的其他文献

CADAL相关文献

Advancing Underwater Vision: A Survey of Deep Learning Models for Underwater Object Recognition and Tracking

读者评论 与其他读者分享你的观点

请选择收藏分类： 新增自定义分类 确定 取消

建议与咨询留下您的常用邮箱和电话号码，以便我们向您反馈解决方案和替代方法

读者评论与其他读者分享你的观点

请选择收藏分类：