Arailym-tleubayeva/sist-kazakh-corpus download history
Arailym-tleubayeva/sist-kazakh-corpus is a sentence similarity dataset on the Hugging Face Hub. In the last 30 days it was downloaded 22 times (6 in the last 7 days), and 206 times in total. It ranks #349,636 among datasets by monthly downloads.
SIST Kazakh Corpus Description SIST Kazakh Corpus is a curated dataset of Kazakh scientific articles collected for research in text similarity detection, plagiarism analysis, and low-resource NLP tasks. The dataset was created to support: Text similarity detection in agglutina
Open Arailym-tleubayeva/sist-kazakh-corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.