whooray/MLDR download history

whooray/MLDR is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 28 times (5 in the last 7 days), and 597 times in total. It ranks #289,218 among datasets by monthly downloads.

Dataset Summary MLDR is a Multilingual Long-Document Retrieval dataset built on Wikipeida, Wudao and mC4, covering 13 typologically diverse languages. Specifically, we sample lengthy articles from Wikipedia, Wudao and mC4 datasets and randomly choose paragraphs from them. Then we use GPT-

Open whooray/MLDR on Hugging Face