idoco/PopVQA download history

idoco/PopVQA is an image text to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 421 times (85 in the last 7 days), and 39,454 times in total. It ranks #37,892 among datasets by monthly downloads.

PopVQA: Popular Entity Visual Question Answering PopVQA is a dataset designed to study the performance gap in vision-language models (VLMs) when answering factual questions about entities presented in images versus text. Paper: https://huggingface.co/papers/2412.14133 Code: https://github

Open idoco/PopVQA on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.