A computer vision pipeline that integrates monocular depth estimation using MiDaS with YOLOv11 object detection with the aim of delivering accurate and reliable weight estimation without the need for using stereo-cameras is presented, making it very well-suited for practical deployment in real-world aquaculture environments.
Abstract
Accurate and automated fish weight estimation is a critical component of modern aquaculture, enabling optimized feeding strategies, improved fish welfare, and effective production planning. A common approach is the use of computer vision techniques to estimate fish mass while the fish are swimming, thereby eliminating the need for manual handling and reducing stress on the fish. This paper presents a computer vision pipeline that integrates monocular depth estimation using MiDaS with YOLOv11 object detection, trained on a real-world underwater dataset of Nile tilapia covering multiple age groups. The new method requires only a monocular camera, eliminates the need for manual fish orientation, and enables fully automatic weight prediction through a CatBoost regressor. Experiments conducted across different fish age groups and tank conditions demonstrate consistent and robust performance. In this context, YOLOv11 achieves an F1-score of 0.93 at an IoU threshold of 0.5. Additionally, the weight estimation model attains a mean absolute error (MAE) below 4 g and an $R^{2}$ value exceeding 0.95. These results suggest that the proposed pipeline approach delivers accurate and reliable weight estimation without the need for using stereo-cameras, making it very well-suited for practical deployment in real-world aquaculture environments.
A novel deep learning framework designed to enhance the accuracy and robustness of fish counting, termed OGLA-Net is proposed and systematically evaluated across three datasets, demonstrating that the proposed counting framework delivers high accuracy across diverse conditions, providing a reliable solution for automat...
Tong-Tong Gu, Zheng-Meng Wu, Da-She Li et al.· Journal of King Saud Univers...· 0 citations
The results suggest that the YOLOv8 object detection model is most effective for abundant and visually distinctive taxa, while performance declines for groups with coarse taxonomic resolution.
Talen Rimmer, Colin Bates, Declan McIntosh et al.· Italian National Conference...· 0 citations
Accurate and efficient fish detection in underwater environments is fundamental to the advancement of automated aquaculture monitoring and selective fishing systems. However, the deployment of existing object detection models in such environments remains challenging due to wavelength-dependent light absorption, suspend...
Van Le, M. Astapova, M. Uzdiaev et al.· Aquaculture Journal· 0 citations
Underwater fish detection is challenged by low light, turbidity, and blue-green color dominance from light attenuation. This study aims to compare six image-enhancement scenarios (baseline, CLAHE, Retinex Ultra Lite, UDP, UDP Super Lite, and Gamma Correction + White Balance) combined with YOLOv11 to evaluate their dete...
Muhammad Iqbal, Indra Jaya, Y. Herdiyeni et al.· Jurnal Teknologi Perikanan d...· 0 citations
Accurate underwater fish perception is a prerequisite for intelligent aquaculture, providing essential support for automated growth monitoring and precision feeding. Nevertheless, underwater imaging is severely affected by light scattering and cluttered backgrounds, making reliable fish perception and image-based non-c...
Mu Ding, Xiao-Hua Huang, Gen Li et al.· Fishes· 0 citations
Real-time detection of marine organisms plays a critical role in underwater ecological monitoring, endangered species protection, and autonomous underwater vehicle (AUV) operations. However, the degraded underwater images with low contrast and detail blur and limited embedded computing resources make it challenging...