zcs15/SVHalluc download history

zcs15/SVHalluc is a visual question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 346 times (233 in the last 7 days), and 1,049 times in total. It ranks #43,945 among datasets by monthly downloads.

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models TL;DR: Speech content is not necessarily visual evidence. SVHalluc is the first benchmark that tests whether audio-visual large language models (AV-LLMs) can distinguish what is said from what i

Open zcs15/SVHalluc on Hugging Face