PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 7, 2026Applied Sciences0 citationsOpen Access

Distracted Driving Behavior Recognition Based on Improved YOLOv8n-Pose and Multi-Feature Fusion

View Full Paper
ZLZhuzhou LiDGDudu GuoZWZhenxun Wei

Key Points

  • The central aim is to enhance the recognition of distracted driving behaviors using an optimized YOLOv8n-Pose model.
  • Optimized YOLOv8n-Pose model with a new detection layer for feature extraction.
  • Incorporated SE attention module for robustness under various lighting conditions.
  • Used a multi-dimensional feature vector based on 12 keypoint coordinates.
  • Employed a BP neural network for classifying the feature vectors.
  • The improved model achieved a mean Average Precision (mAP50) of 93.8% in keypoint detection, surpassing the original model by 6.7%.
  • The BP classification model reached an F1-score of 97.7% for behavior recognition, outperforming traditional classifiers.
  • The system processes images at a speed of 45 frames per second on an NVIDIA RTX 3090TI.

Abstract

Distracted driving is one of the primary causes of road traffic accidents. Behavior recognition technology based on machine vision has emerged as a research hotspot due to its non-contact and high-efficiency nature. To address the challenges of complex lighting conditions in the driver’s cabin, low detection accuracy for small-scale keypoints, and the difficulty in effectively characterizing behavioral features, this paper proposes a distracted driving behavior recognition method based on an improved YOLOv8n-Pose model and multi-feature fusion. First, the original YOLOv8n-Pose model is optimized. A P2 detection layer is added to enhance the feature extraction capabilities for small-scale human keypoints, and the SE attention module is incorporated to improve the model’s robustness under complex lighting conditions. In addition, the loss function is replaced with focal loss to tackle the class imbalance problem, thus forming the YOLOv8n-PSF-Pose keypoint detection network. Subsequently, based on the coordinates of 12 human keypoints extracted by this network, a multi-dimensional feature vector is constructed, which takes joint angles as the core and integrates the relative distances between keypoints and the number of valid keypoints. Finally, a BP neural network is adopted to classify the constructed feature vectors, enabling the accurate recognition of six typical distracted driving behaviors (normal driving, drinking or eating, making phone calls, using mobile phones, operating vehicle infotainment systems, and turning around to fetch items). The experimental results show that the improved YOLOv8n-PSF-Pose model achieves an mAP50 of 93.8% in keypoint detection, which is 6.7 percentage points higher than the original model; the BP classification model based on multi-feature fusion achieves an F1-score of 97.7% in the behavior recognition task, which is significantly better than traditional classifiers such as SVM and random forest, and the image processing speed on the NVIDIA RTX 3090TI reaches a high throughput of 45 FPS. This proves that the proposed method achieves an excellent balance between accuracy and speed. This study provides an effective solution for the real-time and accurate recognition of distracted driving behaviors.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Li et al. (2026) studied this question.

synapsesocial.com/papers/69d49fa9b33cc4c35a228171https://doi.org/10.3390/app16073532
Ask AI
Helpful
Bookmark
Share
View Full Paper