keshan/wit-dataset download history

keshan/wit-dataset is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 35 times (9 in the last 7 days), and 4,452 times in total. It ranks #243,080 among datasets by monthly downloads.

\\nWikipedia-based Image Text (WIT) Dataset is a large multimodal multilingual dataset. WIT is composed of a curated set of 37.6 million entity rich image-text examples with 11.5 million unique images across 108 Wikipedia languages. Its size enables WIT to be used as a pretraining dataset for mult

Open keshan/wit-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.