- Research Article
- 10.17587/prin.15.465-475
Using Large Language Models to Classify Some Vulnerabilities in Program Code
- Sep 09, 2024
- Programmnaya Ingeneria
- V V Shvyrov + 3 more +3
The paper studies the effectiveness of using large language models to detect common types of vulnerabilities in Python program code. In particular, using the technique of low-rank adaptation of (LoRA) models, fine-tuning of the CodeBERT-python model is performed. To train the models, we use the author's dataset, which consists of marked-up program code in Python. The trained models are used to detect and classify potential vulnerabilities. To evaluate the effectiveness of models, the number of false positives, false negatives, true positives and true negatives is determined. Also, accuracy, recall and F1-measures are calculated on a test data set for various configurations of model training macro parameters
Read more