HuggingFaceM4/the_cauldron
Dataset Card for The Cauldron Dataset description The Cauldron is part of the Idefics2 release. It is a massive collection of 50 vision-language datasets (training sets only) that were used for the fine-tuning of the vision-language model Idefics2. Load the dataset To load the dataset, install the library datasets with pip install datasets. Then, from datasets import load_dataset ds = load_dataset("HuggingFaceM4/the_cauldron", "ai2d") to download… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceM4/the_cauldron.
560362k
../
train-00000-of-00012-31f9e352dd92b89f.parquetdownload
train-00001-of-00012-fd5209f92cc41a19.parquetdownload
train-00002-of-00012-92679b2734ac7628.parquetdownload
train-00003-of-00012-a8b56992c61d81ef.parquetdownload
train-00004-of-00012-99a200654bed3175.parquetdownload
train-00005-of-00012-6b1c18f2643c07ae.parquetdownload
train-00006-of-00012-24c23a98e7115d64.parquetdownload
train-00007-of-00012-a0ffeb3aaa88a42b.parquetdownload
train-00008-of-00012-63a22f9c33ecebfa.parquetdownload
train-00009-of-00012-28e0317bd329a609.parquetdownload
train-00010-of-00012-08bd3953e91ca0a1.parquetdownload
train-00011-of-00012-5110ca02b6f250c0.parquetdownload
