amazon-agi/SIFT-50M download history
amazon-agi/SIFT-50M is an audio text to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,470 times (560 in the last 7 days), and 48,243 times in total. It ranks #14,467 among datasets by monthly downloads.
Dataset Card for SIFT-50M SIFT-50M (Speech Instruction Fine-Tuning) is a 50-million-example dataset designed for instruction fine-tuning and pre-training of speech-text large language models (LLMs). It is built from publicly available speech corpora containing a total of 14K hours of spee
Models trained on SIFT-50M
1 models list it as training data.
- kvn420/Tenro_V4.1 0 downloads in 30 days