Skip to content
Open access

Image Classification using DenseNet-121 Based on MediaPipe Face Mesh for Real-Time Drowsiness Detection

Aug 2026 · JOURNAL OF APPLIED INFORMATICS AND COMPUTING · Vol 10, pp. 3704-3718 · 0 citations · 24 references

TL;DR

Findings indicate that the integration of MediaPipe Face Mesh and DenseNet-121 shows meaningful potential for real-time drowsiness monitoring, while also highlighting the importance of subject-independent evaluation and cross-domain generalization for reliable real-world deployment.

Abstract

The high rate of traffic accidents caused by driver drowsiness and microsleep highlights the urgent need for reliable driver monitoring systems. However, conventional Eye Aspect Ratio (EAR) methods often fail due to their high sensitivity to changes in head poses and ambient lighting conditions, while standard Convolutional Neural Network (CNN) models impose heavy computational loads on hardware. This study aims to implement and evaluate a real-time drowsiness detection system by integrating the DenseNet-121 architecture with MediaPipe Face Mesh. The proposed method utilizes MediaPipe Face Mesh to isolate the left and right eye Regions of Interest (ROI) independently, using a proportional padding of 35%, which are then classified using a DenseNet-121 transfer learning model fine-tuned in two stages across its last 30 layers. Evaluation was conducted using a custom dataset of 2,000 source images from five subjects, yielding 3,926 eye-region samples after extraction and quality filtering, assessed using a Subject-Independent Leave-One-Subject-Out (LOSO) cross-validation protocol. Across five folds, the model achieved a mean accuracy of 83.22% (standard deviation 13.22 percentage points) and a mean AUC of 0.879 (standard deviation 0.131), with performance variation across subjects found to correlate with inter-subject differences in eye-closure expressiveness, where the two lowest performing subjects also exhibited the lowest AUC values (0.707 and 0.769). The system achieved an average total latency of 198.00 ms per frame, equivalent to 5.1 FPS. These findings indicate that the integration of MediaPipe Face Mesh and DenseNet-121 shows meaningful potential for real-time drowsiness monitoring, while also highlighting the importance of subject-independent evaluation and cross-domain generalization for reliable real-world deployment.

Read PDF

Similar papers

Conference Aug 2026

Real-Time Edge Computing Framework for Drowsiness Detection in ITS using CLBP

Today’s transportation systems suffer from a high number of accidents caused by drowsy driving, so strong automated detection systems are required for implementation in the Intelligent Transportation Systems (ITS). This paper introduces an analytical framework, divided into three stages to enhance the real-time detecti...

Sumit Sharma, Rakesh Kumar, Meenu Gupta · 0 citations
Open access Aug 2026

Driver Drowsiness Prediction Using CNN-LSTM Model Based on Facial Expression and Eye Movement

Driver fatigue and drowsiness represent primary institutional catalysts for fatal highway traffic anomalies worldwide. This comprehensive investigation introduces an adaptive, multi-task deep learning architecture merging Convolutional Neural Networks and Long Short-Term Memory configurations to dynamically evaluate dr...

Annisa Aprilia Putri Sakri, Angga Rusdinar, G. Mutiara et al. · 0 citations
Open access Aug 2026

Face Mask Detection Using Deep Learning: A Comparative Study of CNN, VGG16 and MobileNetV2 for Real-Time Applications

A comparative study of three deep learning architectures, namely a Custom Convolutional Neural Network (CNN), VGG16, and MobileNetV2, for face mask detection indicates that lightweight transfer learning models offer an effective and practical solution for real-time face mask detection in resource-constrained environmen...

Ruksar Fatima, Shaista Fatima · 0 citations
Conference Aug 2026

A Two-Stage CNN-LSTM Framework for Spatiotemporal Deepfake Video Detection

Now a days identification of convincing synthetic videos created with the help of deepfake is a big challenge. Unfortunately, deepfakes represent a serious threat to the integrity of media, as they can cause individuals to lose trust in the digital content they see. Among all types of deepfakes, face-swap videos are ex...

Pratik Bhosale, Rujul Rajarapollu, Kartikya Durgesh Gawali et al. · 0 citations
Open access Aug 2026

Detection of Face-Swap Based Deepfake Videos Using Hybrid CNN-LSTM Architecture

The development of deepfake technologies due to breakthroughs in AI and deep learning allows producing highly realistic manipulated videos and audio, thus posing a threat to misinformation and digital security. Despite deepfake technology having several legitimate uses, including use in the media industry, its inapprop...

Suraj S. Pawar, Kaustubh R. Saswade, Nikhil R. Mane et al. · 0 citations
Open access Aug 2026

DEEPFAKE FACE DETECTION IN VIDEOS USING OPENCV AND MOBILENETV2

The broad dissemination of altered facial photographs, especially Deepfakes, which are getting harder to identify with traditional techniques, is made possible by the Internet's quick development. While existing methods concentrate on intricate network architectures or geographical domain properties, they sometimes lac...

A. Mohitha, C. B. Jones · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.