pixparse/pdfa-eng-wds download history
pixparse/pdfa-eng-wds is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 7,701 times (1,095 in the last 7 days), and 156,657 times in total. It ranks #3,865 among datasets by monthly downloads.
Dataset Card for PDF Association dataset (PDFA) Dataset Summary PDFA dataset is a document dataset filtered from the SafeDocs corpus, aka CC-MAIN-2021-31-PDF-UNTRUNCATED. The original purpose of that corpus is for comprehensive pdf documents analysis. The purpose of that subset
Models trained on pdfa-eng-wds
17 models list it as training data.
- HuggingFaceM4/idefics2-8b 77K downloads in 30 days
- HuggingFaceM4/idefics2-8b-base 1.4K downloads in 30 days
- HuggingFaceM4/idefics2-8b-chatty 346 downloads in 30 days
- HuggingFaceM4/idefics2-8b-AWQ 104 downloads in 30 days
- mlx-community/idefics2-8b-chatty-4bit 68 downloads in 30 days
- mlx-community/idefics2-8b-4bit 51 downloads in 30 days
- HuggingFaceM4/idefics2-8b-base-AWQ 28 downloads in 30 days
- HuggingFaceM4/idefics2-8b-chatty-AWQ 27 downloads in 30 days
- mlx-community/idefics2-8b-8bit 26 downloads in 30 days
- mlx-community/idefics2-8b-chatty-8bit 24 downloads in 30 days
- huz-relay/idefics2-8b-ocr 23 downloads in 30 days
- Trelis/idefics2-8b-chatty-bf16 17 downloads in 30 days
Spaces using pdfa-eng-wds
- Model Pulse 8 likes
Open pixparse/pdfa-eng-wds on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.