Published June 1, 2026 | Version v1

Predicting Occupational Accident Risk from Textual Data: A Systematic Review of Machine Learning Application

  • 1. Department of Industrial and Systems Engineering, Institut Teknologi Sepuluh Nopember, Indonesia
  • 2. Department of Safety Engineering, Politeknik Perkapalan Negeri Surabaya, Indonesia

Contributors

  • 1. Department of Industrial and Systems Engineering, Institut Teknologi Sepuluh Nopember, Indonesia
  • 2. Department of Safety Engineering, Politeknik Perkapalan Negeri Surabaya, Indonesia

Description

Occupational accident risk remains a persistent concern, especially in high-risk industries. However, the optimal use of textual data in safety analysis is still limited. This study aims to systematically review the application of machine learning (ML) and natural language processing (NLP) in predicting occupational accident risk using textual data. A systematic literature review was conducted using the PRISMA framework and PICOS criteria, with Scopus as the primary database. From 1238 initial articles, 21 were selected for in-depth analysis. The results indicate that algorithms such as Random Forest, Support Vector Machine, and Neural Networks are frequently used, achieving up to 94% accuracy when integrated with NLP techniques. This review highlights the growing potential of ML and NLP in developing proactive and data-driven occupational safety strategies. Further research is recommended to address limitations in data quality, model generalization and validation across various sectors

Notes

Published in Evergreen, Volume 13, Issue 02. Citation formats available via DOI link.

Files

p611-628.pdf

Files (2.2 MB)

Name Size Download all
md5:f010d660322017ab9cb2477a5e592f2f
2.2 MB Preview Download

Additional details

Related works

Is identical to
Journal article: 10.5109/7420072 (DOI)
Is supplemented by
Other: https://citation.crossref.org/?doi=10.5109/7420072 (URL)