google/wit download history
google/wit is a text retrieval dataset on the Hugging Face Hub. In the last 30 days it was downloaded 390 times (102 in the last 7 days), and 12,488 times in total. It ranks #40,131 among datasets by monthly downloads.
Wikipedia-based Image Text (WIT) Dataset is a large multimodal multilingual dataset. WIT is composed of a curated set of 37.6 million entity rich image-text examples with 11.5 million unique images across 108 Wikipedia languages. Its size enables WIT to be used as a pretraining dataset for multimoda
Models trained on wit
2 models list it as training data.
- swap-uniba/LLaVA-NDiNO_pt 11 downloads in 30 days
- GraphicStylz/Stylz – downloads in 30 days