nvidia/describe-anything-dataset download history
nvidia/describe-anything-dataset is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 4,007 times (2,895 in the last 7 days), and 79,308 times in total. It ranks #6,746 among datasets by monthly downloads.
Describe Anything: Detailed Localized Image and Video Captioning NVIDIA, UC Berkeley, UCSF Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui [Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Mod
Models trained on describe-anything-dataset
4 models list it as training data.
- nvidia/DAM-3B-Self-Contained 17K downloads in 30 days
- nvidia/DAM-3B 14.8K downloads in 30 days
- nvidia/DAM-3B-Video 2.2K downloads in 30 days
- randomlysjfsgshzbzf/MultipleModels – downloads in 30 days