MCINext/synthetic-persian-sts download history
MCINext/synthetic-persian-sts is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 91 times (30 in the last 7 days), and 2,950 times in total. It ranks #122,765 among datasets by monthly downloads.
Dataset Summary Synthetic Persian STS is a Persian (Farsi) dataset designed for the Semantic Textual Similarity (STS) task. It is a component of the FaMTEB (Farsi Massive Text Embedding Benchmark). The dataset was synthetically generated using the GPT-4o-mini model, producing sentence pai
Open MCINext/synthetic-persian-sts on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.