Xkev/LLaVA-CoT-100k download history

Xkev/LLaVA-CoT-100k is a visual question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,704 times (353 in the last 7 days), and 44,248 times in total. It ranks #12,867 among datasets by monthly downloads.

Dataset Card for LLaVA-CoT The LLaVA-CoT-100k dataset is introduced in the paper LLaVA-CoT: Let Vision Language Models Reason Step-by-Step. This dataset is designed to enable Vision-Language Models (VLMs) to perform autonomous multistage reasoning, integrating samples from various visual

Models trained on LLaVA-CoT-100k

9 models list it as training data.

Open Xkev/LLaVA-CoT-100k on Hugging Face