Describe Anything: Detailed Localized Image and Video Captioning
344
6 commits
1 linked in READMEs
updated May 1, 2025
NVIDIA, UC Berkeley, UCSF
Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui
[Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Model/Benchmark/Datasets] | [Citation]
Check out the configuration reference at https://huggingface.co/docs/hub/spaces-config-reference
Describe Anything: Detailed Localized Image and Video Captioning
344
6 commits
1 linked in READMEs
updated May 1, 2025
NVIDIA, UC Berkeley, UCSF
Long Lian, Yifan Ding, Yunhao Ge, Sifei Liu, Hanzi Mao, Boyi Li, Marco Pavone, Ming-Yu Liu, Trevor Darrell, Adam Yala, Yin Cui
[Paper] | [Code] | [Project Page] | [Video] | [HuggingFace Demo] | [Model/Benchmark/Datasets] | [Citation]
Check out the configuration reference at https://huggingface.co/docs/hub/spaces-config-reference