Pith. sign in

REVIEW 4 cited by

ChatNVD: Advancing Cybersecurity Vulnerability Assessment with Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.04756 v2 pith:DJM4GK2I submitted 2024-12-06 cs.CR cs.CL

classification cs.CRcs.CL
keywords vulnerabilityassessmentchatnvdmodelssoftwarevulnerabilitiescybersecuritygpt-4o
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The increasing frequency and sophistication of cybersecurity vulnerabilities in software systems underscores the need for more robust and effective vulnerability assessment methods. However, existing approaches often rely on highly technical and abstract frameworks, which hinder understanding and increase the likelihood of exploitation, resulting in severe cyberattacks. In this paper, we introduce ChatNVD, a support tool powered by Large Language Models (LLMs) that leverages the National Vulnerability Database (NVD) to generate accessible, context-rich summaries of software vulnerabilities. We develop three variants of ChatNVD, utilizing three prominent LLMs: GPT-4o Mini by OpenAI, LLaMA 3 by Meta, and Gemini 1.5 Pro by Google. To evaluate their performance, we conduct a comparative evaluation focused on their ability to identify, interpret, and explain software vulnerabilities. Our results demonstrate that GPT-4o Mini outperforms the other models, achieving over 92% accuracy and the lowest error rates, making it the most reliable option for real-world vulnerability assessment.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LLM Embedding-based Attribution (LEA): Quantifying Source Contributions to Generative Model's Response for Vulnerability Analysis

    cs.CR 2025-06 reject novelty 6.0 of 10

    LEA uses rank-based linear dependence of layer-0 hidden states to attribute each response token to query, retrieved context, or internal knowledge, and distinguishes valid from generic retrieval with over 95% accuracy.

  2. On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs

    cs.CR 2024-12 reject novelty 4.0 of 10

    Applying CVSS, DREAD, OWASP, and SSVC to 56 adversarial LLM attacks via three LLM judges yields near-constant factor scores, which the authors take as evidence that these metrics cannot differentiate LLM attacks.

  3. Machine Learning Driven Smishing Detection Framework for Mobile Security

    cs.CR 2024-12 reject novelty 2.0 of 10

    Text normalization with a slang dictionary improves a Naive Bayes smishing classifier from 88.2% to 96.2% accuracy in the authors' experiments, though the evaluation lacks a reproducible dataset and named baselines.

  4. The Future of AI: Exploring the Potential of Large Concept Models

    cs.CL 2025-01 conditional novelty 1.0 of 10

    A survey of grey literature describing Meta's Large Concept Models, their proposed advantages over token-based LLMs, and their speculative applications across many industries.

Pith tools