Published January 1, 2026
| Version v1
Journal article
Open
Lightweight Retrieval-Augmented Generation System For CPU-Only Document Question Answering
Authors/Creators
Description
Retrieval-Augmented Generation (RAG) improves the factual accuracy of Large Language Models by grounding responses in external documents. However, most existing systems rely on dense em-beddings, vector databases, and GPU-based computation, making them unsuitable for low-resource environments. This paper presents a lightweight RAG system designed specifically for CPU-only environments. The system integrates PDF text extraction and Optical Character Recognition (OCR) using PyMuPDF and Tesseract, followed by a keyword-based retrieval mechanism. The retrieved context is then passed to a language model API for response generation. Experimental evaluation demonstrates that the system achieves an accuracy of 83.3% with an average response time of approximately 2.2 seconds. The results highlight that efficient document intelligence systems can be developed without heavy computational requirements
Files
IJSRET_V12_issue2_458.pdf
Files
(548.3 kB)
| Name | Size | Download all |
|---|---|---|
|
md5:aba7e39e58920232f3e9ffd86c5d7c79
|
548.3 kB | Preview Download |
Additional details
Related works
- Has part
- Journal article: https://ijsret.com/wp-content/uploads/IJSRET_V12_issue2_458.pdf (URL)
- Is identical to
- Journal article: https://ijsret.com/2026/04/21/lightweight-retrieval-augmented-generation-system-for-cpu-only-document-question-answering/ (URL)