nvidia/describe-anything-dataset download history

nvidia/describe-anything-dataset is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 4,007 times (2,895 in the last 7 days), and 79,308 times in total. It ranks #6,746 among datasets by monthly downloads.

Describe Anything: Detailed Localized Image and Video Captioning NVIDIA, UC Berkeley, UCSF Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui [Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Mod

Models trained on describe-anything-dataset

4 models list it as training data.

Open nvidia/describe-anything-dataset on Hugging Face