Lo/clip-bert-data
CLIP-BERT training data This data was used to train the CLIP-BERT model first described in this paper. The dataset is based on text and images from MS COCO, SBU Captions, Visual Genome QA and Conceptual Captions. The image features have been extracted using the CLIP model openai/clip-vit-base-patch32 available on Huggingface.
137
Conversations for this repository live on Hugging Face.
Team Ai shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face