tiny-aya-translate/tr-subset-v0.1 download history
tiny-aya-translate/tr-subset-v0.1 is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 430 times (104 in the last 7 days), and 2,621 times in total. It ranks #37,260 among datasets by monthly downloads.
TR Subset v0.1 — Turkish speech 251,118 Turkish audio/text rows (~62 GB, 128 parquet shards). Schema is just text + audio; see the YAML header above. An early-phase Turkish speech collection from the TinyAya data pipeline. It is not part of the v0.3 Stage-2 training corpus — that is t