Team Ai
Datasetpublic

HuggingFaceM4/the_cauldron

Dataset Card for The Cauldron Dataset description The Cauldron is part of the Idefics2 release. It is a massive collection of 50 vision-language datasets (training sets only) that were used for the fine-tuning of the vision-language model Idefics2. Load the dataset To load the dataset, install the library datasets with pip install datasets. Then, from datasets import load_dataset ds = load_dataset("HuggingFaceM4/the_cauldron", "ai2d") to download… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceM4/the_cauldron.

sourceHugging Faceupdated 2y agoView on Hugging Face
560likes362kdownloads
../
filetrain-00000-of-00004-b8bc82a3a0695796.parquet380.4 MBdownload
filetrain-00001-of-00004-c0d095033b5a0a1c.parquet382.2 MBdownload
filetrain-00002-of-00004-6ea8bf17e5f8a869.parquet359.0 MBdownload
filetrain-00003-of-00004-5dcdafe04fc0629d.parquet334.4 MBdownload

HuggingFaceM4/the_cauldron · main · files are served by the source, never re-hosted here