dell-research-harvard/headlines-semantic-similarity download history

dell-research-harvard/headlines-semantic-similarity is a sentence similarity dataset on the Hugging Face Hub. In the last 30 days it was downloaded 785 times (281 in the last 7 days), and 29,439 times in total. It ranks #23,248 among datasets by monthly downloads.

Dataset Card for HEADLINES Dataset Summary HEADLINES is a massive English-language semantic similarity dataset, containing 396,001,930 pairs of different headlines for the same newspaper article, taken from historical U.S. newspapers, covering the period 1920-1989. Lan

Models trained on headlines-semantic-similarity

2 models list it as training data.

Open dell-research-harvard/headlines-semantic-similarity on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.