WSEAS Transactions on Information Science and Applications
Print ISSN: 1790-0832, E-ISSN: 2224-3402
Volume 23, 2026
Optimizing Vision-Based CVS Risk Assessment:
Benchmarking Geometric and Deep Learning Architectures
for Real-Time Applications
Authors: ,
Search Articles
Abstract: This study benchmarks geometric and deep learning architectures for vision-based assessment of behavioral risk factors associated with Computer Vision Syndrome (CVS). Eye blinking, head pose deviation, and body twist were the three behaviors assessed. With accuracies of 0.9034 for blink detection and 0.9037 for head pose estimation, the results demonstrate that Dlib offered the most dependable performance for facial analysis. For head pose analysis, the highest F1-score was 0.7896 at a yaw threshold of 11°, with stable performance across the 9°–11° range. MediaPipe showed greater practical suitability for real-time body twist detection, with processing speeds of 29.98–32.79 FPS, although its maximum accuracy was 0.4668. In contrast, the CNN-based model was constrained by low processing speed, ranging from 4.22 to 7.32 FPS, making it less suitable for immediate deployment. These findings indicate that effective CVS risk assessment requires a balance between detection reliability and computational efficiency. Dlib is more appropriate for accuracy-sensitive facial measurements, whereas MediaPipe offers better practical value for real-time posture and movement monitoring.
Keywords:
Computer Vision Syndrome, CVS Risk Assessment, Behavioral Risk Detection, Dlib, MediaPipe, CNN
Pages: 591-602
DOI: 10.37394/23209.2026.23.50