OmAlve/vaarta-sft-dataset download history

OmAlve/vaarta-sft-dataset is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 15 times (2 in the last 7 days), and 200 times in total. It ranks #451,834 among datasets by monthly downloads.

Vaarta SFT Dataset Multilingual supervised fine-tuning dataset used to train the Vaarta family of Marathi-first language models. ~178K examples in structured messages format — no pre-applied chat template, apply your own at training time. Format Each example has three fields: {

Open OmAlve/vaarta-sft-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.