Kushalkhemka/embedding-cve-nvd-dataset download history

Kushalkhemka/embedding-cve-nvd-dataset is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 26 times (2 in the last 7 days), and 333 times in total. It ranks #306,850 among datasets by monthly downloads.

CVE NVD Embedding Dataset This dataset contains the processed CVE/NVD corpus that was used with the rag_mixedbread pipeline. It bundles: cve_corpus.jsonl (~700 MB): each line is a JSON object with cve_id, title, description, cvss, vendors, and the pre-computed text chunk that feeds the e

Models trained on embedding-cve-nvd-dataset

1 models list it as training data.

Open Kushalkhemka/embedding-cve-nvd-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.