KU-AGI/commonvoice_dualcodec_pretokenize download history
KU-AGI/commonvoice_dualcodec_pretokenize is an automatic speech recognition dataset on the Hugging Face Hub. In the last 30 days it was downloaded 13 times (1 in the last 7 days), and 187 times in total. It ranks #491,332 among datasets by monthly downloads.
Common Voice 17 (English) — DualCodec pre-tokenized DualCodec (12 Hz) pre-tokenized speech for omni speech–vision MLLM training. Source: Mozilla Common Voice 17.0 (en). Contents (train split) Total samples ≈ 1,119,678 (audio–transcript pairs) Total audio
Open KU-AGI/commonvoice_dualcodec_pretokenize on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.