google/wit download history

google/wit is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 390 times (102 in the last 7 days), and 12,488 times in total. It ranks #40,131 among datasets by monthly downloads.

Wikipedia-based Image Text (WIT) Dataset is a large multimodal multilingual dataset. WIT is composed of a curated set of 37.6 million entity rich image-text examples with 11.5 million unique images across 108 Wikipedia languages. Its size enables WIT to be used as a pretraining dataset for multimoda

Models trained on wit

2 models list it as training data.

Open google/wit on Hugging Face