Team Ai
Datasetpublic

friedrichor/ActivityNet_Captions

About ActivityNet Captions contains 20K long-form videos (180s as average length) from YouTube and 100K captions. Most of the videos contain over 3 annotated events. We follow the existing works to concatenate multiple short temporal descriptions into long sentences and evaluate ‘paragraph-to-video’ retrieval on this benchmark. We adopt the official split: Train: 10,009 videos, 10,009 captions (concatenate from 37,421 short captions) Test (Val1): 4,917 videos, 4,917… See the full description on the dataset page: https://huggingface.co/datasets/friedrichor/ActivityNet_Captions.

sourceHugging Faceupdated 1y agoView on Hugging Face
16likes4.3kdownloads
8 commits on main
aaaa81b1y ago

Update README.md

friedrichor
51260b91y ago

Update README.md

friedrichor
17030251y ago

Upload folder using huggingface_hub

friedrichor
d14f17f1y ago

Update README.md

friedrichor
7678bdc2y ago

Update README.md

friedrichor
ce1d5322y ago

Update README.md

friedrichor
cdad2102y ago

Upload folder using huggingface_hub

friedrichor
6d191092y ago

initial commit

friedrichor