<p>Traffic accidents usually result from driver’s inattention, sleepiness, and distraction, posing a substantial danger to worldwide road safety. Advances in computer vision and artificial intelligence (AI) have provided new prospects for designing real-time driver monitoring systems to reduce these dangers. In this paper, we assessed four known deep learning models, MobileNetV2, DenseNet201, NASNetMobile, and VGG19, and offer a unique Hybrid CNN-Transformer architecture reinforced with Efficient Channel Attention (ECA) for multi-class driver activity categorization. The framework defines seven important driving behaviors: Closed Eye, Open Eye, Dangerous Driving, Distracted Driving, Drinking, Yawning, and Safe Driving. Among the baseline models, DenseNet201 (99.40%) and MobileNetV2 (99.31%) achieved the highest validation accuracies. In contrast, the proposed Hybrid CNN-Transformer with ECA attained a near-perfect validation accuracy of 99.72% and further demonstrated flawless generalization with 100% accuracy on the independent test set. Confusion matrix studies further indicate a few misclassifications, verifying the model’s high generalization capacity. By merging CNN-based local feature extraction, attention-driven feature refinement, and Transformer-based global context modeling, the system provides both robustness and efficiency. These findings show the practicality of using the suggested technology in real-time intelligent transportation applications, presenting a viable avenue toward reducing traffic accidents and boosting overall road safety.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Real-time driver activity detection using advanced deep learning models

  • Md. AL Emran,
  • Md. Ariful Islam,
  • Md. Obaydullahn Khan,
  • Md. Jewel Rana,
  • Saida Tasnim Adrita,
  • Md. Ashik Ahmed,
  • Mahmoud M. A. Eid,
  • Ahmed Nabih Zaki Rashed

摘要

Traffic accidents usually result from driver’s inattention, sleepiness, and distraction, posing a substantial danger to worldwide road safety. Advances in computer vision and artificial intelligence (AI) have provided new prospects for designing real-time driver monitoring systems to reduce these dangers. In this paper, we assessed four known deep learning models, MobileNetV2, DenseNet201, NASNetMobile, and VGG19, and offer a unique Hybrid CNN-Transformer architecture reinforced with Efficient Channel Attention (ECA) for multi-class driver activity categorization. The framework defines seven important driving behaviors: Closed Eye, Open Eye, Dangerous Driving, Distracted Driving, Drinking, Yawning, and Safe Driving. Among the baseline models, DenseNet201 (99.40%) and MobileNetV2 (99.31%) achieved the highest validation accuracies. In contrast, the proposed Hybrid CNN-Transformer with ECA attained a near-perfect validation accuracy of 99.72% and further demonstrated flawless generalization with 100% accuracy on the independent test set. Confusion matrix studies further indicate a few misclassifications, verifying the model’s high generalization capacity. By merging CNN-based local feature extraction, attention-driven feature refinement, and Transformer-based global context modeling, the system provides both robustness and efficiency. These findings show the practicality of using the suggested technology in real-time intelligent transportation applications, presenting a viable avenue toward reducing traffic accidents and boosting overall road safety.