Published July 29, 2026
| Version v1
Preprint
Open
MedInsight: Evaluating Retrieval-Augmented Vision-Language Models for Evidence-Grounded Medical Image Understanding
Description
MedInsight evaluates whether retrieval augmentation improves vision-language model performance for medical visual question answering. The study compares a baseline BLIP-2 model with a retrieval-augmented pipeline using CLIP, FAISS, and the ROCOv2 dataset, evaluated on the VQA-RAD benchmark with statistical significance testing. The project includes reproducible code, datasets, and experimental configurations.
Files
MedInsight.pdf
Files
(423.2 kB)
| Name | Size | Download all |
|---|---|---|
|
md5:61c5aefa009df0e3f422c9c59a58e1a5
|
423.2 kB | Preview Download |
Additional details
Software
- Repository URL
- https://github.com/Maryam024/MedInsight
- Programming language
- Python
- Development Status
- Active