tomaarsen/llamaindex-vdr-en-train-preprocessed download history

tomaarsen/llamaindex-vdr-en-train-preprocessed is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 424 times (79 in the last 7 days), and 4,871 times in total. It ranks #37,676 among datasets by monthly downloads.

llamaindex-vdr-en-train-preprocessed This dataset is a preprocessed English subset of llamaindex/vdr-multilingual-train, prepared for training multimodal Sentence Transformer embedding models on document screenshot retrieval. Changes from the original dataset The original llama

Models trained on llamaindex-vdr-en-train-preprocessed

2 models list it as training data.

Open tomaarsen/llamaindex-vdr-en-train-preprocessed on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.