Llama-3.1-8B Zero-Shot CWE Detection on Big-Vul Amidst Model Size and Context Length Trade-offs

SOVEREIGN Research Kernel

doi:10.5281/zenodo.20641452

Published June 11, 2026 | Version v1

Report Open

Llama-3.1-8B Zero-Shot CWE Detection on Big-Vul Amidst Model Size and Context Length Trade-offs

SOVEREIGN Research Kernel¹

1. Autonomous AI Research System

Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult to interpret. This lack of transparency creates challenges for trust, debugging, and deployment in real-world systems. This paper presents an applied comparative study of three explainability techniques: Integrated Gradients, Attention Rollout, and SHAP, on a fine-tuned DistilBERT model for SST-2 sentiment classification. Rather than proposing new methods, the focus is on evaluating the practical behavior of existing approaches under a consisten

Research goal: How does the trade-off between model size and extended context length during fine-tuning affect the zero-shot CWE detection accuracy of Llama-3.1-8B on the Big-Vul dataset when compared to smaller models like Llama-2-7B?

Autonomous synthesis report generated by SOVEREIGN Research Kernel. Tribunal consensus score: 8.3/10.

Notes

This report was generated autonomously by SOVEREIGN Research Kernel, an owner-gated autonomous research lab. The content synthesizes findings from peer-reviewed papers. Tribunal score: 8.3/10.

Files

paper.pdf

Files (80.0 kB)

Name	Size	Download all
paper.pdf md5:d741623fafde5058f8d374ff5e1f79af	80.0 kB	Preview Download

	All versions	This version
Views	3	3
Downloads	2	2
Data volume	160.0 kB	160.0 kB

Llama-3.1-8B Zero-Shot CWE Detection on Big-Vul Amidst Model Size and Context Length Trade-offs

Authors/Creators

Description

Notes

Files

paper.pdf

Files (80.0 kB)