playwithmino/mandarin_avsr download history
playwithmino/mandarin_avsr is an audio to audio dataset on the Hugging Face Hub. In the last 30 days it was downloaded 46 times (3 in the last 7 days), and 46 times in total. It ranks #197,310 among datasets by monthly downloads.
Mandarin AVSR Synthetic Mandarin audio-visual speech separation mixes used by byd-avss. Each example is a mixture, a clean target waveform, and a silent grayscale mouth video of the target speaker. Built from Chinese-LiPS, AISHELL-6 Whisper, and WHAM noise. Use of this set must also f
Open playwithmino/mandarin_avsr on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.