kiyam/ddro-msmarco-doc-dataset-300k download history

kiyam/ddro-msmarco-doc-dataset-300k is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 23 times (7 in the last 7 days), and 1,011 times in total. It ranks #338,096 among datasets by monthly downloads.

DDRO — MS MARCO Top-300K Processed Dataset This dataset contains the preprocessed MS MARCO Top-300K document corpus used to train and evaluate the DDRO generative retrieval models from: Lightweight and Direct Document Relevance Optimization for Generative Information Retrieval (SIGIR 2025

Open kiyam/ddro-msmarco-doc-dataset-300k on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.