Published June 1, 2026
| Version v1
Journal article
Open
Predicting Occupational Accident Risk from Textual Data: A Systematic Review of Machine Learning Application
Authors/Creators
- 1. Department of Industrial and Systems Engineering, Institut Teknologi Sepuluh Nopember, Indonesia
- 2. Department of Safety Engineering, Politeknik Perkapalan Negeri Surabaya, Indonesia
Contributors
Contact person:
- 1. Department of Industrial and Systems Engineering, Institut Teknologi Sepuluh Nopember, Indonesia
- 2. Department of Safety Engineering, Politeknik Perkapalan Negeri Surabaya, Indonesia
Description
Occupational accident risk remains a persistent concern, especially in high-risk industries. However, the optimal use of textual data in safety analysis is still limited. This study aims to systematically review the application of machine learning (ML) and natural language processing (NLP) in predicting occupational accident risk using textual data. A systematic literature review was conducted using the PRISMA framework and PICOS criteria, with Scopus as the primary database. From 1238 initial articles, 21 were selected for in-depth analysis. The results indicate that algorithms such as Random Forest, Support Vector Machine, and Neural Networks are frequently used, achieving up to 94% accuracy when integrated with NLP techniques. This review highlights the growing potential of ML and NLP in developing proactive and data-driven occupational safety strategies. Further research is recommended to address limitations in data quality, model generalization and validation across various sectors
Notes
Files
p611-628.pdf
Files
(2.2 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:f010d660322017ab9cb2477a5e592f2f
|
2.2 MB | Preview Download |
Additional details
Related works
- Is identical to
- Journal article: 10.5109/7420072 (DOI)
- Is supplemented by
- Other: https://citation.crossref.org/?doi=10.5109/7420072 (URL)