vnahata/vaani-audio-image-retrieval download history

vnahata/vaani-audio-image-retrieval is an audio classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 103 times (14 in the last 7 days), and 370 times in total. It ranks #111,475 among datasets by monthly downloads.

Vaani audio–image retrieval (MTEB) Multilingual audio↔image retrieval over 62 Indian languages, derived from Project Vaani (IISc Bangalore / ARTPARK). Vaani records image-prompted speech: a speaker is shown a photograph and describes it aloud in their own language. Each recording is t

Open vnahata/vaani-audio-image-retrieval on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.