Team Ai
Datasetpublic

friedrichor/ActivityNet_Captions

About ActivityNet Captions contains 20K long-form videos (180s as average length) from YouTube and 100K captions. Most of the videos contain over 3 annotated events. We follow the existing works to concatenate multiple short temporal descriptions into long sentences and evaluate ‘paragraph-to-video’ retrieval on this benchmark. We adopt the official split: Train: 10,009 videos, 10,009 captions (concatenate from 37,421 short captions) Test (Val1): 4,917 videos, 4,917… See the full description on the dataset page: https://huggingface.co/datasets/friedrichor/ActivityNet_Captions.

sourceHugging Faceupdated 1y agoView on Hugging Face
16likes4.3kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

Team Ai shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
friedrichor/ActivityNet_Captions · Team Ai